Method and apparatus for processing video signal using affine prediction
Abstract
The present disclosure provides a method for decoding a video signal including a current block based on an affine motion prediction mode (affine mode, AF mode), the method including: checking whether the AF mode is applied to the current block, the AF mode representing a motion prediction mode using an affine motion model; checking whether an AF4 mode is used when the AF mode is applied to the current block, the AF4 mode representing a mode in which a motion vector is predicted using four parameters constituting the affine motion model; generating a motion vector predictor using the four parameters when the AF4 mode is used and generating a motion vector predictor using six parameters constituting the affine motion model when the AF4 mode is not used; and obtaining a motion vector of the current block based on the motion vector predictor.
Claims
exact text as granted — not AI-modifiedWhat is claimed is:
1 . An apparatus for decoding a video signal including a current block based on an affine motion prediction mode (AF mode), comprising:
a processor configured to obtain a merge flag from the video signal, wherein the merge flag represents whether merge mode is applied to the current block, based on that the merge mode not applied to the current block, obtain an affine flag from the video signal based on that a width and a height of the current block is equal to or larger than 16, wherein the affine flag represents whether the AF mode is applied to the current block, and the AF mode represents a motion prediction mode using an affine motion model, obtain an affine parameter flag representing whether 4 parameters or 6 parameters are used for the affine motion model based on that the AF mode is applied to the current block, obtain a motion vector predictor based on the 4 parameters or the 6 parameters being used for the affine motion model, obtain prediction samples for the current block based on the motion vector predictor, obtain residual samples for the current block, reconstruct the current block based on the prediction samples and the residual samples, and filter the reconstructed current block, wherein the affine flag and the affine parameter flag are obtained based on that the merge flag represents the merge mode is not applied to the current block.
2 . The apparatus of claim 1 , wherein the affine flag and the affine parameter flag are defined in a level of a coding unit.
3 . The apparatus of claim 1 , wherein the current block is decoded based on a coding mode other than the AF mode based on that the width and the height of the current block is smaller than the predetermined value.
4 . An apparatus for encoding a video signal including a current block based on an affine motion prediction mode (AF mode), the apparatus comprising:
a processor configured to generate a merge flag representing whether merge mode is applied to the current block, based on that the merge mode not applied to the current block, generate an affine flag based on that a width and a height of the current block is equal to or larger than 16, wherein the affine flag represents whether the AF mode is applied to the current block, and the AF mode represents a motion prediction mode using an affine motion model, generate an affine parameter flag representing whether 4 parameters or 6 parameters are used for the affine motion model based on that the AF mode is applied to the current block, obtain a motion vector predictor based on the 4 parameters or the 6 parameters being used for the affine motion model, generate prediction samples for the current block based on the motion vector predictor, generate residual samples for the current block based on the prediction samples, and perform a transform, a quantization and entropy-encoding for the residual samples, wherein the current block is reconstructed based on the prediction samples and the residual samples, wherein the affine flag and the affine parameter flag are generated based on that the merge flag represents the merge mode is not applied to the current block.
5 . The apparatus of claim 4 , wherein the affine flag and the affine parameter flag are defined in a level of a coding unit.
6 . The apparatus of claim 4 , wherein the current block is encoded based on a coding mode other than the AF mode based on that the width and the height of the current block is smaller than the predetermined value.
7 . A non-transitory computer-readable storage medium storing encoded picture information generated by performing the steps of:
generating a merge flag representing whether merge mode is applied to the current block, based on that the merge mode not applied to the current block, generating an affine flag based on that a width and a height of the current block is equal to or larger than 16, wherein the affine flag represents whether an AF mode is applied to the current block, and the AF mode represents a motion prediction mode using an affine motion model, generating an affine parameter flag representing whether 4 parameters or 6 parameters are used for the affine motion model based on that the AF mode is applied to the current block, obtaining a motion vector predictor based on the 4 parameters or the 6 parameters being used for the affine motion model, generating prediction samples for the current block based on the motion vector predictor, generating residual samples for the current block based on the prediction samples, and performing a transform, a quantization and entropy-encoding for the residual samples, wherein the current block is reconstructed based on the prediction samples and the residual samples, wherein the affine flag and the affine parameter flag are generated based on that the merge flag represents the merge mode is not applied to the current block.
8 . A method of transmitting a bitstream generated by performing the steps of:
generating a merge flag representing whether merge mode is applied to the current block, based on that the merge mode not applied to the current block, generating an affine flag based on that a width and a height of the current block is equal to or larger than 16, wherein the affine flag represents whether an AF mode is applied to the current block, and the AF mode represents a motion prediction mode using an affine motion model, generating an affine parameter flag representing whether 4 parameters or 6 parameters are used for the affine motion model based on that the AF mode is applied to the current block, obtaining a motion vector predictor based on the 4 parameters or the 6 parameters being used for the affine motion model, generating prediction samples for the current block based on the motion vector predictor, generating residual samples for the current block based on the prediction samples, performing a transform, a quantization and entropy-encoding for the residual samples, and transmitting the bitstream including the entropy-encoded residual samples, wherein the current block is reconstructed based on the prediction samples and the residual samples, wherein the affine flag and the affine parameter flag are generated based on that the merge flag represents the merge mode is not applied to the current block.Join the waitlist — get patent alerts
Track US2023353768A1 — get alerts on status changes and closely related new filings.
We store only your email — no account needed. See our privacy policy.