US2018190290A1PendingUtilityA1

Securing audio data

Assignee: GOOGLE LLCPriority: Dec 7, 2016Filed: Feb 28, 2018Published: Jul 5, 2018
Est. expiryDec 7, 2036(~10.4 yrs left)· nominal 20-yr term from priority
G10L 17/22H04L 63/0428G10L 2015/088G10L 15/22G10L 2015/223H04W 12/02G06F 3/165G06F 3/167
45
PatentIndex Score
0
Cited by
0
References
0
Claims

Abstract

Methods, systems, and apparatus, including computer programs encoded on a computer storage medium, for securing audio data. In one aspect, a method includes restricting access by the device to audio information detected by a microphone, receiving data indicating that the device is authorized to access audio information detected by the microphone during a limited period of time, and in response to receiving data indicating that the device is authorized to access audio information detected by the microphone during the limited period of time, providing audio information to the device. The method also includes monitoring audio information detected by the microphone during the limited period of time for the presence of a hotword and after the end of the limited period of time, restricting access by the device to audio information detected by the microphone.

Claims

exact text as granted — not AI-modified
What is claimed is: 
     
         1 . A method comprising:
 obtaining audio information by a microphone of a first device, wherein the audio information includes a first portion, a second portion detected subsequent to detection of the first portion, and a third portion detected subsequent to detection of the second portion;   providing the first portion of the audio information to a second device;   detecting a presence of a hotword in the second portion audio information, wherein the hotword signals a beginning of a voice request;   in response to detecting the presence of the hotword in the second portion of the audio information, blocking the second portion of the audio information from being provided to the second device;   detecting an end of the voice request in the second portion of the audio information; and   in response to detecting the end of the voice request in the second portion of the audio information, providing the third portion of the audio information to the second device.   
     
     
         2 . The method of  claim 1 , wherein detecting a presence of a hotword in the second portion audio information comprises:
 detecting speech corresponding to a predetermined phrase in the second portion audio information.   
     
     
         3 . The method of  claim 1 , wherein detecting an end of the voice request in the second portion of the audio information comprises:
 detecting a pause in speech of a user at the end of audio represented by the second portion of the audio information.   
     
     
         4 . The method of  claim 1 , wherein detecting a pause in speech of a user at the end of audio represented by the second portion of the audio information comprises:
 determining that the pause in speech of the user has a duration that satisfies a criteria based on predetermined length of time.   
     
     
         5 . The method of  claim 1 , wherein blocking the second portion of the audio information from being provided to the second device comprises:
 providing the second portion of the audio information to a voice request server instead of the second device.   
     
     
         6 . The method of  claim 5 , wherein providing the second portion of the audio information to a voice request server instead of the second device comprises:
 encrypting the second portion of the audio information; and   providing the second portion of the audio information in the encrypted form to a processor in the device for the processor to provide to the voice request server instead of the second device.   
     
     
         7 . The method of  claim 1 , wherein providing the third portion of the audio information to the second device comprises:
 detecting that the third portion of the audio information likely does not include a follow-up request; and   in response to detecting that the third portion of the audio information likely does not include a follow-up request, providing the third portion of the audio information to the second device.   
     
     
         8 . The method of  claim 1 , wherein providing the third portion of the audio information to the second device comprises:
 detecting that a predetermined amount of time has passed since the end of the voice request was detected; and   in response to detecting that a predetermined amount of time has passed since the end of the voice request was detected, providing the third portion of the audio information to the second device.   
     
     
         9 . A system comprising:
 one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:   obtaining audio information by a microphone of a first device, wherein the audio information includes a first portion, a second portion detected subsequent to detection of the first portion, and a third portion detected subsequent to detection of the second portion;   providing the first portion of the audio information to a second device;   detecting a presence of a hotword in the second portion audio information, wherein the hotword signals a beginning of a voice request;   in response to detecting the presence of the hotword in the second portion of the audio information, blocking the second portion of the audio information from being provided to the second device;   detecting an end of the voice request in the second portion of the audio information; and   in response to detecting the end of the voice request in the second portion of the audio information, providing the third portion of the audio information to the second device.   
     
     
         10 . The system of  claim 9 , wherein detecting a presence of a hotword in the second portion audio information comprises:
 detecting speech corresponding to a predetermined phrase in the second portion audio information.   
     
     
         11 . The system of  claim 9 , wherein detecting an end of the voice request in the second portion of the audio information comprises:
 detecting a pause in speech of a user at the end of audio represented by the second portion of the audio information.   
     
     
         12 . The system of  claim 9 , wherein detecting a pause in speech of a user at the end of audio represented by the second portion of the audio information comprises:
 determining that the pause in speech of the user has a duration that satisfies a criteria based on predetermined length of time.   
     
     
         13 . The system of  claim 9 , wherein blocking the second portion of the audio information from being provided to the second device comprises:
 providing the second portion of the audio information to a voice request server instead of the second device.   
     
     
         14 . The system of  claim 13 , wherein providing the second portion of the audio information to a voice request server instead of the second device comprises:
 encrypting the second portion of the audio information; and   providing the second portion of the audio information in the encrypted form to a processor in the device for the processor to provide to the voice request server instead of the second device.   
     
     
         15 . The system of  claim 9 , wherein providing the third portion of the audio information to the second device comprises:
 detecting that the third portion of the audio information likely does not include a follow-up request; and   in response to detecting that the third portion of the audio information likely does not include a follow-up request, providing the third portion of the audio information to the second device.   
     
     
         16 . The system of  claim 9 , wherein providing the third portion of the audio information to the second device comprises:
 detecting that a predetermined amount of time has passed since the end of the voice request was detected; and   in response to detecting that a predetermined amount of time has passed since the end of the voice request was detected, providing the third portion of the audio information to the second device.   
     
     
         17 . A non-transitory computer-readable medium storing software comprising instructions executable by one or more computers which, upon such execution, cause the one or more computers to perform operations comprising:
 obtaining audio information by a microphone of a first device, wherein the audio information includes a first portion, a second portion detected subsequent to detection of the first portion, and a third portion detected subsequent to detection of the second portion;   providing the first portion of the audio information to a second device;   detecting a presence of a hotword in the second portion audio information, wherein the hotword signals a beginning of a voice request;   in response to detecting the presence of the hotword in the second portion of the audio information, blocking the second portion of the audio information from being provided to the second device;   detecting an end of the voice request in the second portion of the audio information; and   in response to detecting the end of the voice request in the second portion of the audio information, providing the third portion of the audio information to the second device.   
     
     
         18 . The medium of  claim 17 , wherein detecting a presence of a hotword in the second portion audio information comprises:
 detecting speech corresponding to a predetermined phrase in the second portion audio information.   
     
     
         19 . The medium of  claim 17 , wherein detecting an end of the voice request in the second portion of the audio information comprises:
 detecting a pause in speech of a user at the end of audio represented by the second portion of the audio information.   
     
     
         20 . The medium of  claim 17 , wherein detecting a pause in speech of a user at the end of audio represented by the second portion of the audio information comprises:
 determining that the pause in speech of the user has a duration that satisfies a criteria based on predetermined length of time.

Join the waitlist — get patent alerts

Track US2018190290A1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.