US2014032216A1PendingUtilityA1

Pronunciation Discovery for Spoken Words

Assignee: NUANCE COMMUNICATIONS INCPriority: Sep 11, 2003Filed: Sep 30, 2013Published: Jan 30, 2014
Est. expirySep 11, 2023(expired)· nominal 20-yr term from priority
G10L 15/063G10L 15/187G10L 15/06
50
PatentIndex Score
0
Cited by
0
References
0
Claims

Abstract

A method for a portable device includes receiving a spoken utterance of a word or phrase, generating a plurality of alternative pronunciations of the spoken utterance, scoring one or more pronunciations of the plurality of alternative pronunciations using the spoken utterance, and updating a lexicon with at least one scored pronunciation.

Claims

exact text as granted — not AI-modified
What is claimed is: 
     
         1 . A method comprising:
 receiving a spoken utterance of a word or phrase;   generating a plurality of alternative pronunciations of the spoken utterance;   scoring one or more pronunciations of the plurality of alternative pronunciations using the spoken utterance; and   updating a lexicon with at least one scored pronunciation.   
     
     
         2 . The method of  claim 1 , wherein the at least one scored pronunciation is the highest scoring pronunciation of the scored pronunciations. 
     
     
         3 . The method of  claim 1 , further comprising using a finite-state machine to generate the plurality of alternative pronunciations. 
     
     
         4 . The method of  claim 1 , wherein updating the lexicon includes replacing an existing pronunciation with the at least one scored pronunciation. 
     
     
         5 . The method of  claim 1 , wherein updating the lexicon includes adding a phonetic representation of the at least one scored pronunciation to an existing representation. 
     
     
         6 . The method of  claim 1 , wherein the alternative pronunciations are generated by searching a neighborhood of pronunciations about an initial pronunciation of the spoken utterance. 
     
     
         7 . A system comprising:
 at least one processor; and   a memory device operatively connected to the at least one processor;   wherein, responsive to execution of program instructions accessible to the at least one processor, the at least one processor is configured to:
 receive a spoken utterance of a word or phrase; 
 generate a plurality of alternative pronunciations of the spoken utterance; 
 score one or more pronunciations of the plurality of alternative pronunciations using the spoken utterance; and 
 update a lexicon with at least one scored pronunciation. 
   
     
     
         8 . The system of  claim 7 , wherein the at least one scored pronunciation is the highest scoring pronunciation of the scored pronunciations. 
     
     
         9 . The system of  claim 7 , wherein the at least one processor is configured to use a finite-state machine to generate the plurality of alternative pronunciations. 
     
     
         10 . The system of  claim 7 , wherein the at least one processor is configured to update the lexicon by replacing an existing pronunciation with the at least one scored pronunciation. 
     
     
         11 . The system of  claim 7 , wherein the at least one processor is configured to update the lexicon by adding a phonetic representation of the at least one scored pronunciation to an existing representation. 
     
     
         12 . The system of  claim 7 , wherein the at least one processor is configured to generate the alternative pronunciations by searching a neighborhood of pronunciations about an initial pronunciation of the spoken utterance. 
     
     
         13 . A computer program product encoded in a non-transitory computer-readable medium, which when executed by a computer causes the computer to perform the following operations:
 receiving a spoken utterance of a word or phrase;   generating a plurality of alternative pronunciations of the spoken utterance;   scoring one or more pronunciations of the plurality of alternative pronunciations using the spoken utterance; and   updating a lexicon with at least one scored pronunciation.   
     
     
         14 . The computer program product of  claim 13 , wherein the at least one scored pronunciation is the highest scoring pronunciation of the scored pronunciations. 
     
     
         15 . The computer program product of  claim 13 , wherein the computer uses a finite-state machine to generate the plurality of alternative pronunciations. 
     
     
         16 . The computer program product of  claim 13 , wherein the computer updates the lexicon by replacing an existing pronunciation with the at least one scored pronunciation. 
     
     
         17 . The computer program product of  claim 13 , wherein the computer updates the lexicon by adding a phonetic representation of the at least one scored pronunciation to an existing representation. 
     
     
         18 . The computer program product of  claim 13 , wherein the computer generates the alternative pronunciations by searching a neighborhood of pronunciations about an initial pronunciation of the spoken utterance.

Join the waitlist — get patent alerts

Track US2014032216A1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.