Speech recognition interface for voice actuation of legacy systems
Abstract
Methods and apparatus are disclosed for a technician to access a systems interface to back-end legacy systems by voice input commands to a speech recognition module. Generally, a user logs a computer into a systems interface which permits access to back-end legacy systems. Preferably, the systems interface includes a first server with middleware for managing the protocol interface. Preferably, the systems interface includes a second server for receiving requests and generating legacy transactions. After the computer is logged-on, a request for voice input is made. A speech recognition module is launched or otherwise activated. The user inputs voice commands that are processed to convert them to commands and text that can be recognized by the client software. The client software formats the requests and forwards them to the systems interface in order to retrieve the requested information.
Claims
exact text as granted — not AI-modifiedWe claim:
1 . A method comprising:
presenting, via a display, a plurality of operations available to a user for interacting with a legacy system; receiving a multimodal request from the user to perform a transaction, wherein:
the multimodal request has a speech portion that is recognized by an input device;
the multimodal request has a second portion comprising a cursor input; and
the multimodal request corresponds to an operation in the plurality of operations;
recognizing the speech portion via a vocabulary specific to the plurality of operations presented; translating the multimodal request to a format compatible with the legacy system, to yield a translated multimedia request; and submitting the translated multimodal request to the legacy system.
2 . The method of claim 1 , wherein communications with the legacy system occur over a wireless communications network.
3 . The method of claim 1 , wherein communications with the legacy system occur over a wireline communications network.
4 . The method of claim 1 , wherein the plurality of operations are presented via a graphical user interface on the display.
5 . The method of claim 1 , wherein a plurality of vocabularies are used for the plurality of operations, each vocabulary in the plurality of vocabularies corresponding to an operation in the plurality of operations.
6 . The method of claim 1 , wherein the speech portion corresponds to a legacy operation of the legacy system.
7 . The method of claim 6 , wherein the legacy operation was a cursory input operation.
8 . A system comprising:
a processor; and a computer-readable storage medium having instructions stored which, when executed by the processor, result in the processor performing operations comprising:
presenting, via a display, a plurality of operations available to a user for interacting with a legacy system;
receiving a multimodal request from the user to perform a transaction, wherein:
the multimodal request has a speech portion that is recognized by an input device;
the multimodal request has a second portion comprising a cursor input; and
the multimodal request corresponds to an operation in the plurality of operations;
recognizing the speech portion via a vocabulary specific to the plurality of operations presented;
translating the multimodal request to a format compatible with the legacy system, to yield a translated multimedia request; and
submitting the translated multimodal request to the legacy system.
9 . The system of claim 8 , wherein communications with the legacy system occur over a wireless communications network.
10 . The system of claim 8 , wherein communications with the legacy system occur over a wireline communications network.
11 . The system of claim 8 , wherein the plurality of operations are presented via a graphical user interface on the display.
12 . The system of claim 8 , wherein a plurality of vocabularies are used for the plurality of operations, each vocabulary in the plurality of vocabularies corresponding to an operation in the plurality of operations.
13 . The system of claim 8 , wherein the speech portion corresponds to a legacy operation of the legacy system.
14 . The system of claim 13 , wherein the legacy operation was a cursory input operation.
15 . A computer-readable storage device having instructions stored which, when executed by a computing device, result in the computing device performing operations comprising:
presenting, via a display, a plurality of operations available to a user for interacting with a legacy system; receiving a multimodal request from the user to perform a transaction, wherein:
the multimodal request has a speech portion that is recognized by an input device;
the multimodal request has a second portion comprising a cursor input; and
the multimodal request corresponds to an operation in the plurality of operations;
recognizing the speech portion via a vocabulary specific to the plurality of operations presented; translating the multimodal request to a format compatible with the legacy system, to yield a translated multimedia request; and submitting the translated multimodal request to the legacy system.
16 . The computer-readable storage device of claim 15 , wherein communications with the legacy system occur over a wireless communications network.
17 . The computer-readable storage device of claim 15 , wherein communications with the legacy system occur over a wireline communications network.
18 . The computer-readable storage device of claim 15 , wherein the plurality of operations are presented via a graphical user interface on the display.
19 . The computer-readable storage device of claim 15 , wherein a plurality of vocabularies are used for the plurality of operations, each vocabulary in the plurality of vocabularies corresponding to an operation in the plurality of operations.
20 . The computer-readable storage device of claim 15 , wherein the speech portion corresponds to a legacy operation of the legacy system.Join the waitlist — get patent alerts
Track US2016026433A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.