US11935557B2ActiveUtilityA1

Techniques for detecting and processing domain-specific terminology

Assignee: HARMAN INT INDPriority: Feb 1, 2021Filed: Feb 1, 2021Granted: Mar 19, 2024
Est. expiryFeb 1, 2041(~14.5 yrs left)· nominal 20-yr term from priority
G10L 25/78G10L 15/08G10L 21/0208G06F 3/165G06F 40/237G06F 40/279
53
PatentIndex Score
0
Cited by
13
References
13
Claims

Abstract

Various embodiments set forth systems and techniques for explaining domain-specific terms detected in a media content stream. The techniques include detecting a speech portion included in an audio signal; determining that the speech portion comprises a domain-specific term; determining an explanatory phrase associated with the domain-specific term; and integrating the explanatory phrase associated with the domain-specific term into playback of the audio signal.

Claims

exact text as granted — not AI-modified
What is claimed is: 
     
       1. A computer-implemented method for explaining domain-specific terms detected in a media content stream, the computer-implemented method comprising:
 detecting a speech portion included in an audio signal; 
 determining that the speech portion comprises a domain-specific term; 
 disabling audio pass-through of the audio signal after determining that the speech portion comprises the domain-specific term; 
 storing, in an audio buffer, a portion of the audio signal after the domain-specific term; 
 determining an explanatory phrase associated with the domain-specific term; 
 replacing the domain-specific term with the explanatory phrase in playback of the audio signal; and 
 playing back the explanatory phrase; 
 after playing back the explanatory phrase, playing back the buffered portion of the audio signal at an increased playback speed; and 
 in response to determining that the buffered portion of the audio signal has caught up with the audio signal:
 stopping playback of the buffered portion of the audio signal, and 
 enabling audio pass-through of the audio signal. 
 
 
     
     
       2. The computer-implemented method of  claim 1 , wherein the explanatory phrase is based on one or more contextual cues associated with the speech portion. 
     
     
       3. The computer-implemented method of  claim 2 , wherein the one or more contextual cues include one or more words before or after the speech portion. 
     
     
       4. The computer-implemented method of  claim 1 , wherein the increased playback speed is determined based on at least one of:
 a network speed, or 
 a length of the buffered portion of the audio signal. 
 
     
     
       5. The computer-implemented method of  claim 1 , wherein the step of replacing the domain-specific term with the explanatory phrase in the playback of the audio signal comprises:
 generating an acoustic representation of the explanatory phrase; and 
 replacing the domain-specific term with the acoustic representation. 
 
     
     
       6. The computer-implemented method of  claim 1 , further comprising enabling voice cancellation or noise cancellation during the step of playing back the explanatory phrase. 
     
     
       7. The computer-implemented method of  claim 1 , wherein the domain-specific term is associated with domains of knowledge known to a user and stored in a user profile of the user. 
     
     
       8. A system, comprising:
 a memory storing one or more software applications; and 
 a processor that, when executing the one or more software applications, is configured to perform the steps of:
 detecting a speech portion included in an audio signal; 
 determining that the speech portion comprises a domain-specific term; 
 disabling audio pass-through of the audio signal in response to determining that the speech portion comprises the domain-specific term; 
 storing, in an audio buffer, a portion of the audio signal after the domain-specific term; 
 determining an explanatory phrase associated with the domain-specific term; 
 replacing the domain-specific term with the explanatory phrase in playback of the audio signal; 
 playing back the explanatory phrase; 
 after playing back the explanatory phrase, playing back the buffered portion of the audio signal at an increased playback speed; and 
 in response to determining that the buffered portion of the audio signal has caught up with the audio signal:
 stopping playback of the buffered portion of the audio signal, and 
 enabling audio pass-through of the audio signal. 
 
 
 
     
     
       9. The system of  claim 8 , wherein:
 the explanatory phrase is generated based on one or more search results retrieved from one or more data stores; and 
 one or more weights are applied to the one or more search results based on reliability of the one or more data stores. 
 
     
     
       10. The system of  claim 8 , wherein the increased playback speed is determined based on at least one of:
 a network speed, or 
 a length of the buffered portion of the audio signal. 
 
     
     
       11. One or more non-transitory computer-readable media storing instructions that, when executed by one or more processors, cause the one or more processors to perform the steps of:
 detecting a speech portion included in an audio signal; 
 determining that the speech portion comprises a domain-specific term; 
 disabling audio pass-through of the audio signal after determining that the speech portion comprises the domain-specific term; 
 storing, in an audio buffer, a portion of the audio signal after the domain-specific term; 
 determining an explanatory phrase associated with the domain-specific term; 
 replacing the domain-specific term with the explanatory phrase in playback of the audio signal; and 
 playing back the explanatory phrase; 
 after playing back the explanatory phrase, playing back the buffered portion of the audio signal at an increased playback speed; and 
 in response to determining that the buffered portion of the audio signal has caught up with the audio signal:
 stopping playback of the buffered portion of the audio signal, and 
 enabling audio pass-through of the audio signal. 
 
 
     
     
       12. The one or more non-transitory computer-readable media of  claim 11 , further storing instructions that, when executed by the one or more processors, cause the one or more processors to perform the steps of:
 generating an acoustic representation of the explanatory phrase; and 
 replacing the domain-specific term with the acoustic representation. 
 
     
     
       13. The one or more non-transitory computer-readable media of  claim 11 , further storing instructions that, when executed by the one or more processors, cause the one or more processors to perform the steps of:
 enabling voice cancellation or noise cancellation during the step of playing back the explanatory phrase.

Join the waitlist — get patent alerts

Track US11935557B2 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.