Computer-implemented methods and systems for modeling and recognition of speech
Abstract
In accordance with the present invention, computer implemented methods and systems are provided for representing and modeling the temporal structure of audio signals. In response to receiving a signal, a time-to-frequency domain transformation on at least a portion of the received signal to generate a frequency domain representation is performed. The time-to-frequency domain transformation converts the signal from a time domain representation to the frequency domain representation. A frequency domain linear prediction (FDLP) is performed on the frequency domain representation to estimate a temporal envelope of the frequency domain representation. Based on the temporal envelope, one or more speech features are generated.
Claims
exact text as granted — not AI-modified1 . A method of extracting speech features from signals for use in performing automatic speech recognition, the method comprising: receiving a signal; performing a time-to-frequency domain transformation on at least a portion of the received signal to generate a frequency domain representation; performing a frequency domain linear prediction on the frequency domain representation to estimate a temporal envelope of the frequency domain representation; and generating at least one speech feature based at least in part on the temporal envelope.
Join the waitlist — get patent alerts
Track US2009271182A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.