US2020211581A1PendingUtilityA1

Speech enhancement using clustering of cues

Assignee: KARDOME TECH LTDPriority: Oct 19, 2017Filed: Dec 23, 2019Published: Jul 2, 2020
Est. expiryOct 19, 2037(~11.2 yrs left)· nominal 20-yr term from priority
Inventors:Alon Slapak
G10L 2021/02166G10L 21/0232G10L 21/0224G10L 25/90
50
PatentIndex Score
0
Cited by
0
References
0
Claims

Abstract

A method for speech enhancement, the method may include receiving or generating sound samples that represent sound signals that were received during a given time period by an array of microphones; frequency transforming the sound samples to provide frequency-transformed samples; clustering the frequency-transformed samples to speakers to provide speaker related clusters, wherein the clustering is based on (i) spatial cues related to the received sound signals and (ii) acoustic cues related to the speakers; determining a relative transfer function for each speaker of the speakers to provide speakers related relative transfer functions; applying a multiple input multiple output (MIMO) beamforming operation on the speakers related relative transfer functions to provide beamformed signals; and inverse-frequency transforming the beamformed signals to provide speech signals.

Claims

exact text as granted — not AI-modified
We claim: 
     
         1 . A method for speech enhancement, the method comprises:
 receiving or generating sound samples that represent sound signals that were received during a given time period by an array of microphones;   frequency transforming the sound samples to provide frequency-transformed samples;   clustering the frequency-transformed samples to speakers to provide speaker related clusters, wherein the clustering is based on (i) spatial cues related to the received sound signals and (ii) acoustic cues related to the speakers;   determining a relative transfer function for each speaker of the speakers to provide speakers related relative transfer functions;   applying a multiple input multiple output (MIMO) beamforming operation on the speakers related relative transfer functions to provide beamformed signals;   inverse-frequency transforming the beamformed signals to provide speech signals.

Join the waitlist — get patent alerts

Track US2020211581A1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.