Responding to Human Spoken Audio Based on User Input
Abstract
Systems and methods for responding to human spoken are provided herein. Exemplary methods may include receiving audio input for generating a speech signal using at least one microphone communicatively coupled to an intelligent assistant device. The method may also include transmitting the audio input from the intelligent assistant device to a natural language processor, the audio input having been converted from speech to a text query. The method may further include processing the text query using artificial intelligence (AI) logic, determining an Application Programming Interface (API) from a plurality of APIs for processing the text query, and transmitting a response from the API to the intelligent assistant device or another device communicatively coupled to the intelligent assistant device for output.
Claims
exact text as granted — not AI-modifiedWhat is claimed is:
1 . A system comprising:
an intelligent assistant device comprising a processor which executes logic to perform operations comprising:
receiving audio input for generating a speech signal using at least one microphone communicatively coupled to the intelligent assistant device; and
a natural language processor communicatively coupled with the intelligent assistant device that executes logic to perform operations comprising:
receiving the audio input from the intelligent assistant device;
converting the audio input from speech to a text query;
processing the text query using artificial intelligence (AI) logic;
determining an Application Programming Interface (API) from a plurality of APIs for processing the text query; and
transmitting a response from the API to the intelligent assistant device or another device communicatively coupled to the intelligent assistant device for output.
2 . The system of claim 1 , wherein the natural language processor further uses machine learning to analyze a string of text.
3 . The system of claim 1 , wherein the natural language processor uses a neural network to analyze the string of text.
4 . The system of claim 1 , wherein the natural language processor interprets and learns from patterns and behaviors of a user, attributing data to the patterns and behaviors such that a response to a future command from the user can be automatically generated by the intelligent assistant device.
5 . The system of claim 1 , wherein the intelligent assistant device acts as a base station connected to at least one enabled device such that the audio input received by the intelligent assistant device is used to adjust an operation of the connected at least one enabled device.
6 . The system of claim 5 , wherein the at least one enabled device receives data to transmit to or from the intelligent assistant device.
7 . The system of claim 5 , wherein the at least one enabled device is connected to the intelligent assistant device via Bluetooth.
8 . The system of claim 5 , wherein the at least one enabled device comprises a smartphone comprising:
at least one microphone for receiving audio commands; at least one user input interface; a mobile application for processing audio and user input commands; and a natural language processor for performing automatic speech recognition of the audio commands.
9 . The system of claim 5 , wherein the at least one enabled device comprises at least one smart home device.
10 . The system of claim 5 , wherein the at least one smart home device receives commands from a general server connected to the intelligent assistant device.
11 . The system of claim 1 , wherein the intelligent assistant device utilizes digital signal processing to separate background noise in the audio input.
12 . The system of claim 1 , wherein the intelligent assistant device includes indicators that provide interactive feedback.
13 . A method, comprising:
receiving audio input for generating a speech signal using at least one microphone communicatively coupled to an intelligent assistant device; transmitting the audio input from the intelligent assistant device to a natural language processor; converting the audio input from speech to a text query using the natural language processor; processing the text query using artificial intelligence (AI) logic using the natural language processor; determining an Application Programming Interface (API) from a plurality of APIs for processing the text query using the natural language processor; and transmitting a response from the API to the intelligent assistant device or another device communicatively coupled to the intelligent assistant device for output using the natural language processor.
14 . The method of claim 13 , further comprising processing the text query using machine learning to analyze a string of text using the natural language processor.
15 . The method of claim 13 , further comprising processing the text query using a neural network to analyze the string of text using the natural language processor.
16 . The method of claim 13 , further comprising connecting the intelligent assistant device to at least one enabled device such that the audio input received by the intelligent assistant device is used to adjust an operation of the connected at least one enabled device.
17 . The method of claim 16 , wherein the at least one enabled device receives data to transmit to or from the intelligent assistant device
18 . The method of claim 16 , wherein the at least one enabled device comprises a smartphone comprising:
at least one microphone for receiving audio commands; at least one user input interface; a mobile application for processing audio and user input commands; and a natural language processor for performing automatic speech recognition of the audio commands.
19 . The method of claim 16 , wherein the at least one enabled device comprises at least one smart home device that receives commands from a general server connected to the intelligent assistant device.
20 . An interactive device, comprising:
at least one microphone; at least one speaker; and a processor that executes logic stored in memory to perform operations comprising:
receiving audio input for generating a speech signal using the at least one microphone;
transmitting the audio input from the device to a natural language processor;
receiving a response to the audio input from a server; and
outputting the response.Join the waitlist — get patent alerts
Track US2017046124A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.