US12520027B1ActiveUtility

Framing visual content during capture

Assignee: GOPRO INCPriority: Nov 22, 2023Filed: Nov 22, 2023Granted: Jan 6, 2026
Est. expiryNov 22, 2043(~17.3 yrs left)· nominal 20-yr term from priority
Inventors:BELHAKIMI AMINE
H04N 23/698H04N 23/55H04N 23/632H04N 23/611H04N 23/58
35
PatentIndex Score
0
Cited by
10
References
20
Claims

Abstract

An image capture device may capture visual content of a video. During the capture of the video, a user may perform a framing interaction with the image capture device. The framing interaction with the image capture device may provide information/instruction on how the visual content being captured by the image capture device should be framed. The framing interaction with the image capture device may assign a direction of the visual content for framing. The framing interaction performed during the capture of the visual content may be used to frame the visual content after capture.

Claims

exact text as granted — not AI-modified
What is claimed is: 
     
         1 . An image capture device for framing videos, the image capture device comprising:
 a housing;   an optical element carried by the housing and configured to guide light within a field of view to an image sensor;   the image sensor carried by the housing and configured to generate a visual output signal conveying visual information based on light that becomes incident thereon, the visual information defining visual content; and   one or more physical processors carried by the housing and configured by machine-readable instructions to:
 capture the visual content during a capture duration to generate a video, the video having a progress length; 
 during a moment within the capture duration, detect a framing interaction by a user with the image capture device to frame the visual content, the framing interaction by the user with the image capture device to frame the visual content including the user speaking a voice command or the user making a hand gesture during the capture duration, the framing interaction by the user with the image capture device received from a direction from the image capture device, the direction of the framing interaction including a direction from which the voice command was received by the image capture device or a direction of the hand gestured depicted within the visual content, the moment within the capture duration corresponding to a moment within the progress length of the video; 
   wherein:
 the framing interaction during the capture duration causes performance of object tracking for video framing and determines a direction from which the object tracking is started; 
 a location of depiction of the user within the visual content at the moment within the progress length of the video is determined based on the direction from which the framing interaction by the user was received by the image capture device during the capture duration; 
 object tracking is performed to track changes in locations of the depiction of the user within the visual content over a duration after the moment within the progress length of the video; and 
 framing of the video at the moment within the progress length and over the duration after the moment within the progress length is determined based on the locations of the depiction of the user at the moment within the progress length and over the duration after the moment within the progress length, the framing of the video defining positioning of a viewing window for the visual content within the video. 
   
     
     
         2 . The image capture device of  claim 1 , wherein determination of the location of depiction of the user within the visual content at the moment within the progress length of the video based on the direction from which the framing interaction by the user was received by the image capture device during the capture duration includes:
 performing face detection in the direction from which the voice command was received by the image capture device during the capture duration; or   performing person detection in the direction of the hand gestured depicted within the visual content.   
     
     
         3 . An image capture device for framing videos, the image capture device comprising:
 a housing;   an optical element carried by the housing and configured to guide light within a field of view to an image sensor;   the image sensor carried by the housing and configured to generate a visual output signal conveying visual information based on light that becomes incident thereon, the visual information defining visual content; and   one or more physical processors carried by the housing and configured by machine-readable instructions to:
 capture the visual content during a capture duration to generate a video, the video having a progress length; 
 during a moment within the capture duration, detect a framing interaction by a user with the image capture device to frame the visual content, the framing interaction by the user with the image capture device received from a direction from the image capture device, the moment within the capture duration corresponding to a moment within the progress length of the video; 
   wherein:
 the framing interaction during the capture duration causes performance of object tracking for video framing and determines a direction from which the object tracking is started; 
 a location of depiction of the user within the visual content at the moment within the progress length of the video is determined based on the direction from which the framing interaction by the user was received by the image capture device during the capture duration; 
 object tracking is performed to track changes in locations of the depiction of the user within the visual content over a duration after the moment within the progress length of the video; and 
 framing of the video at the moment within the progress length and over the duration after the moment within the progress length is determined based on the locations of the depiction of the user at the moment within the progress length and over the duration after the moment within the progress length, the framing of the video defining positioning of a viewing window for the visual content within the video. 
   
     
     
         4 . The image capture device of  claim 3 , wherein the framing interaction with the image capture device to frame the visual content includes the user speaking a voice command during the capture duration, and the direction from which the framing interaction by the user was received by the image capture device includes a direction from which the voice command was received by the image capture device. 
     
     
         5 . The image capture device of  claim 4 , wherein determination of the location of depiction of the user within the visual content at the moment within the progress length of the video based on the direction from which the framing interaction by the user was received by the image capture device during the capture duration includes performing face detection in the direction from which the voice command was received by the image capture device during the capture duration. 
     
     
         6 . The image capture device of  claim 3 , wherein the framing interaction with the image capture device to frame the visual content includes the user making a hand gesture during the capture duration, and the direction from which the framing interaction by the user was received by the image capture device includes a direction of the hand gestured depicted within the visual content. 
     
     
         7 . The image capture device of  claim 6 , wherein determination of the location of depiction of the user within the visual content at the moment within the progress length of the video based on the direction from which the framing interaction by the user was received by the image capture device during the capture duration includes performing person detection in the direction of the hand gestured depicted within the visual content. 
     
     
         8 . The image capture device of  claim 3 , wherein determination of the framing of the video at the moment within the progress length includes determination of a viewing direction for the viewing window at the moment within the progress length. 
     
     
         9 . The image capture device of  claim 8 , wherein the determination of the framing of the video at the moment within the progress length further includes determination of a viewing size for the viewing window at the moment within the progress length. 
     
     
         10 . The image capture device of  claim 3 , wherein the visual content is captured to generate a spherical video. 
     
     
         11 . The image capture device of  claim 3 , wherein:
 the locations of depictions of the user within the visual content are determined as bounding boxes positioned within the visual content; and   the viewing window is positioned to include the bounding boxes.   
     
     
         12 . A method for framing videos, the method performed by an image capture device including an optical element, an image sensor, and one or more processors, the optical element configured to guide light within a field of view to an image sensor, the image sensor configured to generate a visual output signal conveying visual information based on light that becomes incident thereon, the visual information defining visual content, the method comprising:
 capturing the visual content during a capture duration to generate a video, the video having a progress length;   during a moment within the capture duration, detecting a framing interaction by a user with the image capture device to frame the visual content, the framing interaction by the user with the image capture device received from a direction from the image capture device, the moment within the capture duration corresponding to a moment within the progress length of the video;   wherein:
 the framing interaction during the capture duration causes performance of object tracking for video framing and determines a direction from which the object tracking is started; 
 a location of depiction of the user within the visual content at the moment within the progress length of the video is determined based on the direction from which the framing interaction by the user was received by the image capture device during the capture duration; 
 object tracking is performed to track changes in locations of the depiction of the user within the visual content over a duration after the moment within the progress length of the video; and 
 framing of the video at the moment within the progress length and over the duration after the moment within the progress length is determined based on the locations of the depiction of the user at the moment within the progress length and over the duration after the moment within the progress length, the framing of the video defining positioning of a viewing window for the visual content within the video. 
   
     
     
         13 . The method of  claim 12 , wherein the framing interaction with the image capture device to frame the visual content includes the user speaking a voice command during the capture duration, and the direction from which the framing interaction by the user was received by the image capture device includes a direction from which the voice command was received by the image capture device. 
     
     
         14 . The method of  claim 13 , wherein determination of the location of depiction of the user within the visual content at the moment within the progress length of the video based on the direction from which the framing interaction by the user was received by the image capture device during the capture duration includes performing face detection in the direction from which the voice command was received by the image capture device during the capture duration. 
     
     
         15 . The method of  claim 12 , wherein the framing interaction with the image capture device to frame the visual content includes the user making a hand gesture during the capture duration, and the direction from which the framing interaction by the user was received by the image capture device includes a direction of the hand gestured depicted within the visual content. 
     
     
         16 . The method of  claim 15 , wherein determination of the location of depiction of the user within the visual content at the moment within the progress length of the video based on the direction from which the framing interaction by the user was received by the image capture device during the capture duration includes performing person detection in the direction of the hand gestured depicted within the visual content. 
     
     
         17 . The method of  claim 12 , wherein determination of the framing of the video at the moment within the progress length includes determination of a viewing direction for the viewing window at the moment within the progress length. 
     
     
         18 . The method of  claim 17 , wherein the determination of the framing of the video at the moment within the progress length further includes determination of a viewing size for the viewing window at the moment within the progress length. 
     
     
         19 . The method of  claim 12 , wherein the visual content is captured to generate a spherical video. 
     
     
         20 . The method of  claim 12 , wherein:
 the locations of depictions of the user within the visual content are determined as bounding boxes positioned within the visual content; and   the viewing window is positioned to include the bounding boxes.

Join the waitlist — get patent alerts

Track US12520027B1 — get alerts on status changes and closely related new filings.

We store only your email — no account needed. See our privacy policy.