Systems and methods for automatically adjusting audio based on gaze point
Abstract
Embodiments provide methods and systems for adjusting audio output based on eye tracking input. In some embodiments, a memory stores data defining a boundary based on a coordinate system. The boundary corresponds to a display element of displayed content. An input receives data indicating coordinates of a gaze point location of a user viewing the displayed content. A processor compares the received coordinates of the gaze point location to the boundary corresponding to the display element to determine whether the gaze point location is inside the boundary corresponding to the display element. In response to determining that the gaze point location is inside the boundary corresponding to the display element, the processor adjusts an audio setting of the displayed content.
Claims
exact text as granted — not AI-modified1 - 30 . (canceled)
31 . A system for adjusting an image based on eye tracking input, the system comprising:
a memory storing data defining a boundary based on a coordinate system, the boundary corresponding to a first display element of displayed content; an input device configured to receive data indicating coordinates of a gaze point location of a user viewing the displayed content; and a processor configured to:
compare the received coordinates of the gaze point location to the boundary corresponding to the first display element to determine whether the gaze point location is outside the boundary corresponding to the first display element; and
adjust an image of the displayed content in response to determining that the gaze point location is outside the boundary corresponding to the first display element.
32 . The system of claim 31 , wherein the processor configured to adjust the image is further configured to:
bring a second display element of the displayed content into focus; and bring the first display element of the displayed content out of focus.
33 . The system of claim 32 , wherein the second display element is brought into focus subsequent to a predefined duration set by a user.
34 . The system of claim 33 , wherein the predefined duration set by the user is a period of uninterrupted time for which the user's gaze is directed to the second display element.
35 . The system of claim 31 , wherein the processor is further configured to generate for display closed captioning in response to determining that the gaze point location is outside the boundary corresponding to the first display element.
36 . The system of claim 31 , wherein the processor is further configured to:
in response to determining that the gaze point location is outside the boundary corresponding to the first display element, adjust an audio setting corresponding to the displayed content such that audio corresponding to the first display element is at a volume level is lower than a volume level of audio corresponding a second display element of the displayed content.
37 . The system of claim 31 , wherein the processor is further configured to:
select a first audio track associated with a second display element from a plurality of audio tracks associated with the displayed content in response to determining that the gaze point location is outside the boundary corresponding to the first display element; and present the selected first audio track instead of a second of the plurality of audio tracks that is associated with the first display element.
38 . The system of claim 37 , wherein the processor is further configured to adjust at least one additional audio track associated with the displayed content.
39 . The system of claim 31 , wherein the processor is further configured to select an audio track to accompany the content from a plurality of audio tracks associated with the displayed content.
40 . The system of claim 31 , the system further comprising an eye tracker configured to:
determine a gaze point of the user; determine coordinates of a location on the display that the gaze point corresponds to; and transmit data indicating the coordinates of the gaze point location on the display to the input device.
41 . A method for adjusting an image based on eye tracking input, the method comprising:
storing data defining a boundary based on a coordinate system, the boundary corresponding to a first display element of displayed content; receiving data indicating coordinates of a gaze point location of a user viewing the displayed content; comparing, using control circuitry, the received coordinates of the gaze point location to the boundary corresponding to the first display element to determine whether the gaze point location is outside the boundary corresponding to the first display element; and adjusting, using control circuitry, an image of the displayed content in response to determining that the gaze point location is outside the boundary corresponding to the first display element.
42 . The method of claim 41 , wherein the adjusting comprises:
bringing a second display element of the displayed content into focus; and bringing the first display element of the displayed content out of focus.
43 . The method of claim 42 , wherein the second display element is brought into focus subsequent to a predefined duration set by a user.
44 . The method of claim 43 , wherein the predefined duration set by the user is a period of uninterrupted time for which the user's gaze is directed to the second display element.
45 . The method of claim 41 , further comprising:
generating for display closed captioning in response to determining that the gaze point location is outside the boundary corresponding to the first display element.
46 . The method of claim 41 , further comprising:
in response to determining that the gaze point location is outside the boundary corresponding to the first display element, adjusting an audio setting corresponding to the displayed content such that audio corresponding to the first display element is at a volume level is lower than a volume level of audio corresponding a second display element of the displayed content.
47 . The method of claim 41 , further comprising:
selecting a first audio track associated with a second display element from a plurality of audio tracks associated with the displayed content in response to determining that the gaze point location is outside the boundary corresponding to the first display element; and presenting the selected first audio track instead of a second of the plurality of audio tracks that is associated with the first display element.
48 . The method of claim 47 , further comprising adjusting at least one additional audio track associated with the displayed content.
49 . The method of claim 41 , further comprising selecting an audio track to accompany the content from a plurality of audio tracks associated with the displayed content.
50 . The method of claim 41 , further comprising:
determining a gaze point of the user; determining coordinates of a location on the display that the gaze point corresponds to; and transmitting data indicating the coordinates of the gaze point location on the display.Join the waitlist — get patent alerts
Track US2014375558A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.