US2023110205A1PendingUtilityA1

Alternate natural language input generation

Assignee: AMAZON TECH INCPriority: Dec 4, 2019Filed: Sep 1, 2022Published: Apr 13, 2023
Est. expiryDec 4, 2039(~13.4 yrs left)· nominal 20-yr term from priority
G10L 15/22G10L 15/142G10L 15/1815G10L 15/30G10L 2015/088G10L 2015/223G10L 15/197G10L 25/78G06F 40/30
60
PatentIndex Score
0
Cited by
0
References
0
Claims

Abstract

Techniques for handling errors during processing of natural language inputs are described. A system may process a natural language input to generate an ASR hypothesis or NLU hypothesis. The system may use more than one data searching technique (e.g., deep neural network searching, convolutional neural network searching, etc.) to generate an alternate ASR hypothesis or NLU hypothesis, depending on the type of hypothesis input for alternate hypothesis processing.

Claims

exact text as granted — not AI-modified
1 .- 20 . (canceled) 
     
     
         21 . A computer-implemented method, comprising:
 receiving audio data representing a spoken natural language input;   performing speech processing to determine a first speech processing hypothesis corresponding to a first interpretation of the spoken natural language input;   determining a likelihood that further processing of the first speech processing hypothesis will result in a processing error;   based at least in part on the likelihood, determining a second speech processing hypothesis corresponding to a second interpretation of the spoken natural language input different from the first interpretation;   determining the second speech processing hypothesis corresponds to an action; and   causing performance of the action.   
     
     
         22 . The computer-implemented method of  claim 21 , further comprising:
 determining the first speech processing hypothesis corresponding to a first entity; and   determining the second speech processing hypothesis corresponding to a second entity different from the first entity.   
     
     
         23 . The computer-implemented method of  claim 21 , wherein:
 performing speech processing comprises performing automatic speech recognition (ASR) using the audio data to determine a first ASR hypothesis, wherein the first speech processing hypothesis comprises the first ASR hypothesis; and   determining the second speech processing hypothesis comprises determining a second ASR hypothesis different from the first ASR hypothesis.   
     
     
         24 . The computer-implemented method of  claim 21 , wherein:
 performing speech processing comprises performing natural language understanding (NLU) using first data representing the audio data to determine a first NLU hypothesis, wherein the first speech processing hypothesis comprises the first NLU hypothesis; and   determining the second speech processing hypothesis comprises determining a second NLU hypothesis different from the first NLU hypothesis.   
     
     
         25 . The computer-implemented method of  claim 21 , further comprising:
 sending, from a speech processing component, to a first component, first data corresponding to the first speech processing hypothesis; and   processing the first data using the first component to determine the likelihood.   
     
     
         26 . The computer-implemented method of  claim 25 , further comprising:
 processing, by the first component, the first data with respect to stored data to determine the likelihood, the stored data corresponding to at least one prior speech processing hypothesis.   
     
     
         27 . The computer-implemented method of  claim 21 , further comprising:
 processing first data corresponding to the first speech processing hypothesis with respect to stored data to determine a prior speech processing hypothesis that resulted in a system performing a correct action; and   selecting the prior speech processing hypothesis as the second speech processing hypothesis.   
     
     
         28 . The computer-implemented method of  claim 27 , further comprising:
 determining a user profile corresponding to the spoken natural language input; and   determining the stored data based at least in part on the user profile.   
     
     
         29 . The computer-implemented method of  claim 21 , further comprising:
 determining that a confidence value associated with the first speech processing hypothesis fails to satisfy a confidence threshold.   
     
     
         30 . The computer-implemented method of  claim 29 , wherein determination of the second speech processing hypothesis occurs after determining that the confidence value associated with the first speech processing hypothesis fails to satisfy the confidence threshold. 
     
     
         31 . A system comprising:
 at least one processor; and   at least one memory comprising instructions that, when executed by the at least one processor, cause the system to:
 receive audio data representing a spoken natural language input; 
 perform speech processing to determine a first speech processing hypothesis corresponding to a first interpretation of the spoken natural language input; 
 determine a likelihood that further processing of the first speech processing hypothesis will result in a processing error; 
 based at least in part on the likelihood, determine a second speech processing hypothesis corresponding to a second interpretation of the spoken natural language input different from the first interpretation; 
 determine the second speech processing hypothesis corresponds to an action; and 
 cause performance of the action. 
   
     
     
         32 . The system of  claim 31 , wherein the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to:
 determine the first speech processing hypothesis corresponding to a first entity; and   determine the second speech processing hypothesis corresponding to a second entity different from the first entity.   
     
     
         33 . The system of  claim 31 , wherein:
 the instructions that cause the system to perform speech processing comprise instructions that, when executed by the at least one processor, cause the system to perform automatic speech recognition (ASR) using the audio data to determine a first ASR hypothesis, wherein the first speech processing hypothesis comprises the first ASR hypothesis; and   the instructions that cause the system to determine the second speech processing hypothesis comprise instructions that, when executed by the at least one processor, cause the system to determine a second ASR hypothesis different from the first ASR hypothesis.   
     
     
         34 . The system of  claim 31 , wherein:
 the instructions that cause the system to perform speech processing comprise instructions that, when executed by the at least one processor, cause the system to perform natural language understanding (NLU) using first data representing the audio data to determine a first NLU hypothesis, wherein the first speech processing hypothesis comprises the first NLU hypothesis; and   the instructions that cause the system to determine the second speech processing hypothesis comprise instructions that, when executed by the at least one processor, cause the system to determine a second NLU hypothesis different from the first NLU hypothesis.   
     
     
         35 . The system of  claim 31 , wherein the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to:
 send, from a speech processing component, to a first component, first data corresponding to the first speech processing hypothesis; and   process the first data using the first component to determine the likelihood.   
     
     
         36 . The system of  claim 35 , wherein the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to:
 process, by the first component, the first data with respect to stored data to determine the likelihood, the stored data corresponding to at least one prior speech processing hypothesis.   
     
     
         37 . The system of  claim 31 , wherein the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to:
 process first data corresponding to the first speech processing hypothesis with respect to stored data to determine a prior speech processing hypothesis that resulted in a system performing a correct action; and   select the prior speech processing hypothesis as the second speech processing hypothesis.   
     
     
         38 . The system of  claim 37 , wherein the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to:
 determine a user profile corresponding to the spoken natural language input; and   determine the stored data based at least in part on the user profile.   
     
     
         39 . The system of  claim 31 , wherein the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to:
 determine that a confidence value associated with the first speech processing hypothesis fails to satisfy a confidence threshold.   
     
     
         40 . The system of  claim 39 , wherein determination of the second speech processing hypothesis occurs after determination that the confidence value associated with the first speech processing hypothesis fails to satisfy the confidence threshold.

Join the waitlist — get patent alerts

Track US2023110205A1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.