System and method for providing remote automatic speech...

Data processing: speech signal processing – linguistics – language – Speech signal processing – Application

Reexamination Certificate

Rate now

  [ 0.00 ] – not rated yet Voters 0   Comments 0

Details

C704S275000, C704S243000, C707S793000

Reexamination Certificate

active

06366886

ABSTRACT:

TECHNICAL FIELD
This invention relates to speech recognition in general and, more particularly, provides a way of providing remotely-accessible automatic speech recognition services via a packet network.
BACKGROUND OF THE INVENTION
Techniques for accomplishing automatic speech recognition (ASR) are well known. Among known ASR techniques are those that use grammars. A grammar is a representation of the language or phrases expected to be used or spoken in a given context. In one sense, then, ASR grammars typically constrain the speech recognizer to a vocabulary that is a subset of the universe of potentially-spoken words; and grammars may include sub-grammars. An ASR grammar rule can then be used to represent the set of “phrases” or combinations of words from one or more grammars or subgrammars that may be expected in a given context. “Grammar” may also refer generally to a statistical language model (where a model represents phrases), such as those used in language understanding systems.
Products and services that utilize some form of automatic speech recognition (“ASR”) methodology have been recently introduced commercially. For example, AT&T has developed a grammar-based ASR engine called WATSON that enables development of complex ASR services. Desirable attributes of complex ASR services that would utilize such ASR technology include high accuracy in recognition; robustness to enable recognition where speakers have differing accents or dialects, and/or in the presence of background noise; ability to handle large vocabularies; and natural language understanding. In order to achieve these attributes for complex ASR services, ASR techniques and engines typically require computer-based systems having significant processing capability in order to achieve the desired speech recognition capability. Processing capability as used herein refers to processor speed, memory, disk space, as well as access to application databases. Such requirements have restricted the development of complex ASR services that are available at one's desktop, because the processing requirements exceed the capabilities of most desktop systems, which are typically based on personal computer (PC) technology.
Packet networks are general-purpose data networks which are well-suited for sending stored data of various types, including speech or audio. The Internet, the largest and most renowned of the existing packet networks, connects over 4 million computers in some 140 countries. The Internet's global and exponential growth is common knowledge today.
Typically, one accesses a packet network, such as the Internet, through a client software program executing on a computer, such as a PC, and so packet networks are inherently client/server oriented. One way of accessing information over a packet network is through use of a Web browser (such as the Netscape Navigator, available from Netscape Communications, Inc., and the Internet Explorer, available from Microsoft Corp.) which enables a client to interact with Web servers. Web servers and the information available therein are typically identified and addressed through a Uniform Resource Locator (URL)-compatible address. URL addressing is widely used in Internet and intranet applications and is well known to those skilled in the art (an “intranet” is a packet network modeled in functionality based upon the Internet and is used, e.g., by companies locally or internally).
What is desired is a way of enabling ASR services that may be made available to users at a location, such as at their desktop, that is remote from the system hosting the ASR engine.
SUMMARY OF THE INVENTION
A system and method of operating an automatic speech recognition service using a client-server architecture is used to make ASR services accessible at a client location remote from the location of the main ASR engine. In accordance with the present invention, using client-server communications over a packet network, such as the Internet, the ASR server receives a grammar from the client, receives information representing speech from the client, performs speech recognition, and returns information based upon the recognized speech to the client. Alternative embodiments of the present invention include a variety of ways to obtain access to the desired grammar, use of compression or feature extraction as a processing step at the ASR client prior to transferring speech information to the ASR server, staging a dialogue between client and server, and operating a form-filling service.


REFERENCES:
patent: 5231691 (1993-07-01), Yasuda
patent: 5425128 (1995-06-01), Morrison
patent: 5475792 (1995-12-01), Stanford et al.
patent: 5524169 (1996-06-01), Cohen et al.
patent: 5548729 (1996-08-01), Akiyoshi et al.
patent: 5553119 (1996-09-01), McAllister et al.
patent: 5623605 (1997-04-01), Keshav et al.
patent: 5682478 (1997-10-01), Watson et al.
patent: 5732219 (1998-03-01), Blumer et al.
patent: 5745754 (1998-04-01), Lagarde et al.
patent: 5745874 (1998-04-01), Neely
patent: 5752232 (1998-05-01), Basore et al.
patent: 5890123 (1999-03-01), Brown et al.
patent: 6078886 (2000-06-01), Dragosh et al.
patent: 0607615 (1994-07-01), None
patent: 0854418 (1998-07-01), None
patent: 07222248 (1995-08-01), None
“Integrating Speech Technology with Voice Response Units Systems”, IBM Technical Disclosure Bulletin, vol. 38, No. 10, Oct. 1, 1995, pp. 215-216, XP000540471.
Carlson, S., et al.: “Application of Speech Recognition Technology To ITS Advanced Traveloer Information Systems”, Pacific Rim Transtech Conference Vehicle Navigation and Information Systems Conference Proceedings, Washington, Jul. 30-Aug. 2, 1995, No. CONF. 6, Jul. 30, 1995, pp. 118-125, XP000641143, IEEE, ¶ II.
“Client-Server Model for Speech Recognition”, IBM Technical Disclosure Bulletin, vol. 36, No. 3, Mar. 1, 1993, pp. 25-26, XP000354688.

LandOfFree

Say what you really think

Search LandOfFree.com for the USA inventors and patents. Rate them and share your experience with other people.

Rating

System and method for providing remote automatic speech... does not yet have a rating. At this time, there are no reviews or comments for this patent.

If you have personal experience with System and method for providing remote automatic speech..., we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and System and method for providing remote automatic speech... will most certainly appreciate the feedback.

Rate now

     

Profile ID: LFUS-PAI-O-2923565

  Search
All data on this website is collected from public sources. Our data reflects the most accurate information available at the time of publication.