Voice prompts for use in speech-to-speech translation system
Abstract
Techniques for employing improved prompts in a speech-to-speech translation system are disclosed. By way of example, a technique for use in indicating a dialogue turn in an automated speech-to-speech translation system comprises the following steps/operations. One or more text-based scripts are obtained. The one or more text-based scripts are synthesizable into one or more voice prompts. At least one of the one or more voice prompts is synthesized for playback from at least one of the one or more text-based scripts, the at least one synthesized voice prompt comprising an audible message in a language understandable to a speaker interacting with the speech-to-speech translation system, the audible message indicating a dialogue turn in the automated speech-to-speech translation system.
Claims
exact text as granted — not AI-modified1 . A method for use in indicating a dialogue turn in an automated speech-to-speech translation system, comprising the steps of:
obtaining one or more text-based scripts, the one or more text-based scripts being synthesizable into one or more voice prompts; and synthesizing for playback at least one of the one or more voice prompts from at least one of the one or more text-based scripts, the at least one synthesized voice prompt comprising an audible message in a language understandable to a speaker interacting with the speech-to-speech translation system, the audible message indicating a dialogue turn in the automated speech-to-speech translation system.
2 . The method of claim 1 , further comprising the step of detecting a language spoken by a speaker interacting with the speech-to-speech translation system such that a voice prompt in the detected language is synthesized for playback to the speaker.
3 . The method of claim 2 , wherein an initial voice prompt is synthesized for playback in a default language until the actual language of the speaker is detected.
4 . The method of claim 1 , further comprising the step of displaying the at least one voice prompt synthesized for playback.
5 . The method of claim 1 , further comprising the step of recognizing speech uttered by the speaker interacting with the speech-to-speech translation system.
6 . The method of claim 5 , further comprising the step of recognizing speech uttered by a system user of the speech-to-speech translation system.
7 . The method of claim 6 , wherein at least a portion of the speech uttered by the speaker or the system user is translated from one language to another language.
8 . The method of claim 7 , wherein at least a portion of the translated speech is displayed.
9 . A method of providing an interface for use in an automated speech-to-speech translation system, the translation system being operated by a system user and interacted with by a speaker, the method comprising the steps of:
the system user enabling a microphone of the translation system via the interface; outputting at least one previously-generated voice prompt to the speaker, the at least one voice prompt comprising an audible message in a language understandable to the speaker, the audible message indicating a turn in a dialogue between the system user and the speaker; and the speaker, once prompted, uttering speech into the microphone, the uttered speech being translated by the translation system.
10 . The method of claim 9 , further comprising the step of displaying text in a first field of the interface representing speech uttered by the system user.
11 . The method of claim 10 , further comprising the step of displaying text in a second field of the interface representing speech uttered by the speaker.
12 . Apparatus for use in indicating a dialogue turn in an automated speech-to-speech translation system, comprising:
a memory; and at least one processor coupled to the memory and operative to: (i) obtain one or more text-based scripts, the one or more text-based scripts being synthesizable into one or more voice prompts, and (ii) synthesize for playback at least one of the one or more voice prompts from at least one of the one or more text-based scripts, the at least one synthesized voice prompt comprising an audible message in a language understandable to a speaker interacting with the speech-to-speech translation system, the audible message indicating a dialogue turn in the automated speech-to-speech translation system.
13 . The apparatus of claim 12 , wherein the at least one processor is further operative to detect a language spoken by a speaker interacting with the speech-to-speech translation system such that a voice prompt in the detected language is synthesized for playback to the speaker.
14 . The apparatus of claim 13 , wherein an initial voice prompt is synthesized for playback in a default language until the actual language of the speaker is detected.
15 . The apparatus of claim 12 , wherein the at least one processor is further operative to display the at least one voice prompt synthesized for playback.
16 . The apparatus of claim 12 , wherein the at least one processor is further operative to recognize speech uttered by the speaker interacting with the speech-to-speech translation system.
17 . The apparatus of claim 16 , wherein the at least one processor is further operative to recognize speech uttered by a system user of the speech-to-speech translation system.
18 . The apparatus of claim 17 , wherein at least a portion of the speech uttered by the speaker or the system user is translated from one language to another language.
19 . The apparatus of claim 18 , wherein at least a portion of the translated speech is displayed.
20 . An interface for use in an automated speech-to-speech translation system, the translation system being operated by a system user and interacted with by a speaker, the interface comprising:
a first field for use by the system user to enable a microphone of the translation system; a second field for use by the system user for at least one of displaying speech uttered by the system user and displaying translated speech uttered by the speaker; and a third field for use by the speaker for at least one of displaying speech uttered by the speaker and displaying translated speech uttered by the system user; wherein the translation system outputs at least one previously-generated voice prompt to the speaker, the at least one voice prompt comprising an audible message in a language understandable to the speaker, the audible message indicating a turn in a dialogue between the system user and the speaker, and the speaker, once prompted, uttering speech into the microphone, the uttered speech being translated by the translation system.
21 . The interface of claim 20 , further comprising a fourth field for use by the system user to enable a microphone of the translation system such that speech uttered by the system user is captured by the translation system.
22 . An article of manufacture for use in indicating a dialogue turn in an automated speech-to-speech translation system, comprising a machine readable medium containing one or more programs which when executed implement the steps of:
obtaining one or more text-based scripts, the one or more text-based scripts being synthesizable into one or more voice prompts; and synthesizing for playback at least one of the one or more voice prompts from at least one of the one or more text-based scripts, the at least one synthesized voice prompt comprising an audible message in a language understandable to a speaker interacting with the speech-to-speech translation system, the audible message indicating a dialogue turn in the automated speech-to-speech translation system.Join the waitlist — get patent alerts
Track US2006253272A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.