User-directed navigation of multimedia search results
Abstract
A method and apparatus for timed tagging of content is featured. The method and apparatus can include the steps of, or structure for, obtaining at least one keyword tag associated with discrete media content; generating a timed segment index of discrete media content, the timed segment index identifying content segments of the discrete media content and corresponding timing boundaries of the content segments; searching the timed segment index for a match to the at least one keyword tag, the match corresponding to at least one of the content segments identified in the segment index; and generating a timed tag index that includes the at least one keyword tag and the timing boundaries corresponding to the least one content segment of the discrete media content containing the match.
Claims
exact text as granted — not AI-modified1 - 15 . (canceled)
16 . A system for presenting search results for media content, the system comprising a media processor configured to:
receive discrete media content; derive timing offsets from the discrete media content using one or more automated media processing techniques; and present a search result that enables a user to arbitrarily select and commence playback of the discrete media content at any of the content segments of the discrete media content.
17 . The system of claim 16 , wherein the media processor is further configured to:
present the search result including transcriptions of one or more of the content segments of the discrete media content, each of the transcriptions being mapped to a timing offset of a corresponding content segment; receive a user selection of one of the transcriptions presented in the search result; and cause playback of the discrete media content at a timing offset of the corresponding content segment mapped to the selected one of the transcriptions.
18 . The system of claim 17 , wherein each of the transcriptions is derived from the discrete media content using one or more automated media processing techniques or obtained from closed caption data associated with the discrete media content.
19 . The system of claim 17 , wherein the search result further comprises a user actuated display element that enables the user to navigate from an offset of one content segment to another content segment within the discrete media content in response to user actuation of the element.
20 . The system of claim 16 , wherein the search result further comprises a user actuated display element that enables the user to navigate from an offset of one content segment to another content segment within the discrete media content in response to user actuation of the element.
21 . The system of claim 20 , wherein the media processor is further configured to:
obtain timing offsets corresponding to each of the content segments within the discrete media content; in response to an indication of user actuation of the display element, determine a playback offset associated with the discrete media content in playback; compare the playback offset with the timing offsets corresponding to each of the content segments to determine which of the content segments is presently in playback; and cause playback of the discrete media content to continue at an offset that is prior to or subsequent to the offset of the content segment presently in playback.
22 . The system of claim 16 , wherein one or more of the content segments comprise word segments, audio speech segments, video segments, non-speech audio segments, or marker segments.
23 . The system of claim 16 , wherein one or more of the content segments comprise audio corresponding to an individual word, audio corresponding to a phrase, audio corresponding to a sentence, audio corresponding to a paragraph, audio corresponding to a story, audio corresponding to a topic, audio within a range of volume levels, audio of an identified speaker, audio during a speaker turn, audio associated with a speaker emotion, audio of non-speech sounds, audio separated by sound gaps, audio separated by markers embedded within the media content or audio corresponding to a named entity.
24 . The system of claim 16 , wherein one or more of the content segments comprise video of individual scenes, watermarks, recognized objects, recognized faces, overlay text or video separated by markers embedded within the media content.
25 . The system of claim 17 , wherein each of the transcriptions is associated with a confidence level, and the media processor is further configured to:
present the search result including the transcriptions of the one or more of the content segments of the discrete media content, such that any transcription that is associated with a confidence level that fails to satisfy a predefined threshold is displayed with one or more predefined symbols.
26 . An apparatus for presenting search results for media content, comprising:
means for presenting a search result that enables a user to arbitrarily select and commence playback of the discrete media content at any of the content segments of the discrete media content using timing offsets derived from the discrete media content using one or more automated media processing techniques.Join the waitlist — get patent alerts
Track US2009222442A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.