CN110636298B - Unified constraints for Merge affine mode and non-Merge affine mode - Google Patents
Unified constraints for Merge affine mode and non-Merge affine mode Download PDFInfo
- Publication number
- CN110636298B CN110636298B CN201910544642.0A CN201910544642A CN110636298B CN 110636298 B CN110636298 B CN 110636298B CN 201910544642 A CN201910544642 A CN 201910544642A CN 110636298 B CN110636298 B CN 110636298B
- Authority
- CN
- China
- Prior art keywords
- affine mode
- video block
- merge affine
- merge
- video
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Active
Links
Images
Classifications
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04N—PICTORIAL COMMUNICATION, e.g. TELEVISION
- H04N19/00—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals
- H04N19/50—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using predictive coding
- H04N19/503—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using predictive coding involving temporal prediction
- H04N19/51—Motion estimation or motion compensation
- H04N19/55—Motion estimation with spatial constraints, e.g. at image or region borders
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04N—PICTORIAL COMMUNICATION, e.g. TELEVISION
- H04N19/00—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals
- H04N19/10—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding
- H04N19/102—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the element, parameter or selection affected or controlled by the adaptive coding
- H04N19/103—Selection of coding mode or of prediction mode
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04N—PICTORIAL COMMUNICATION, e.g. TELEVISION
- H04N19/00—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals
- H04N19/10—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding
- H04N19/102—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the element, parameter or selection affected or controlled by the adaptive coding
- H04N19/103—Selection of coding mode or of prediction mode
- H04N19/109—Selection of coding mode or of prediction mode among a plurality of temporal predictive coding modes
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04N—PICTORIAL COMMUNICATION, e.g. TELEVISION
- H04N19/00—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals
- H04N19/10—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding
- H04N19/134—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the element, parameter or criterion affecting or controlling the adaptive coding
- H04N19/136—Incoming video signal characteristics or properties
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04N—PICTORIAL COMMUNICATION, e.g. TELEVISION
- H04N19/00—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals
- H04N19/10—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding
- H04N19/169—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the coding unit, i.e. the structural portion or semantic portion of the video signal being the object or the subject of the adaptive coding
- H04N19/17—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the coding unit, i.e. the structural portion or semantic portion of the video signal being the object or the subject of the adaptive coding the unit being an image region, e.g. an object
- H04N19/176—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the coding unit, i.e. the structural portion or semantic portion of the video signal being the object or the subject of the adaptive coding the unit being an image region, e.g. an object the region being a block, e.g. a macroblock
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04N—PICTORIAL COMMUNICATION, e.g. TELEVISION
- H04N19/00—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals
- H04N19/10—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding
- H04N19/169—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the coding unit, i.e. the structural portion or semantic portion of the video signal being the object or the subject of the adaptive coding
- H04N19/184—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the coding unit, i.e. the structural portion or semantic portion of the video signal being the object or the subject of the adaptive coding the unit being bits, e.g. of the compressed video stream
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04N—PICTORIAL COMMUNICATION, e.g. TELEVISION
- H04N19/00—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals
- H04N19/10—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding
- H04N19/169—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the coding unit, i.e. the structural portion or semantic portion of the video signal being the object or the subject of the adaptive coding
- H04N19/186—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using adaptive coding characterised by the coding unit, i.e. the structural portion or semantic portion of the video signal being the object or the subject of the adaptive coding the unit being a colour or a chrominance component
-
- H—ELECTRICITY
- H04—ELECTRIC COMMUNICATION TECHNIQUE
- H04N—PICTORIAL COMMUNICATION, e.g. TELEVISION
- H04N19/00—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals
- H04N19/50—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using predictive coding
- H04N19/503—Methods or arrangements for coding, decoding, compressing or decompressing digital video signals using predictive coding involving temporal prediction
- H04N19/51—Motion estimation or motion compensation
- H04N19/513—Processing of motion vectors
Abstract
Devices, systems, and methods for sub-block based prediction are described. In a representative aspect, a method for video processing includes: the method includes determining block size constraints, making a determination as to whether Merge affine mode and non-Merge affine mode are allowed for video blocks in a video frame based on the block size constraints, and generating a bitstream representation of the video blocks based on making the determination.
Description
Cross Reference to Related Applications
The present application is required to claim in time the priority and benefit of international patent application No. PCT/CN2018/092118 filed on 21.6.6.2018 according to applicable patent laws and/or according to the rules of the paris convention. The entire disclosure of International patent application No. PCT/CN2018/092118 is incorporated by reference herein as part of the disclosure of the present application.
Technical Field
This patent document relates generally to image and video coding techniques.
Background
Motion compensation is a technique in video processing that predicts frames in a video given previous and/or future frames by taking into account the motion of the camera and/or objects in the video. Motion compensation may be used for encoding and decoding of video data for video compression.
Disclosure of Invention
Devices, systems, and methods related to subblock-based prediction for image and video encoding are described.
In representative aspects, the disclosed techniques may be used to provide a method for video processing. The method includes determining a block size constraint, making a determination as to whether Merge affine mode and non-Merge affine mode are allowed for a video block in a video frame based on the block size constraint, and generating a bitstream representation of the video block based on the making the determination.
In a representative aspect, the disclosed techniques may be used to provide another method for video processing. The method includes determining a block size constraint, making a determination as to whether Merge affine mode and non-Merge affine mode are allowed for a video block in a video frame based on the block size constraint, and generating a video block from a bitstream representation of the video block based on making the determination.
In yet another representative aspect, the above-described methods are implemented in the form of processor executable code and stored in a computer readable program medium.
In yet another representative aspect, an apparatus configured or operable to perform the above-described method is disclosed. The apparatus may include a processor programmed to implement the method.
In yet another representative aspect, a video decoder device may implement the methods described herein.
The above and other aspects and features of the disclosed technology are described in more detail in the accompanying drawings, the description and the claims.
Drawings
Fig. 1 shows an example of sub-block based prediction.
Fig. 2 shows an example of a simplified affine motion model.
Fig. 3 shows an example of an affine Motion Vector Field (MVF) for each sub-block.
Fig. 4 shows an example of Motion Vector Prediction (MVP) for the AF _ INTER affine motion mode.
Fig. 5A and 5B show example candidates for the AF _ MERGE affine motion mode.
FIG. 6 shows a block diagram for use in JEM with 4: 2: example of sub-blocks of different components of the 0 format.
Fig. 7A shows a flow diagram of an example method for video processing.
Fig. 7B illustrates another flow diagram of an example method for video processing.
FIG. 8 is a block diagram illustrating an example of an architecture of a computer system or other control device that may be used to implement various portions of the presently disclosed technology.
FIG. 9 illustrates a block diagram of an example embodiment of a device that may be used to implement portions of the presently disclosed technology.
Detailed Description
Due to the increasing demand for higher resolution video, video encoding methods and techniques are ubiquitous in modern technology. Video codecs typically include electronic circuits or software that compress or decompress digital video and are continually being improved to provide higher coding efficiency. The video codec converts uncompressed video into a compressed format or converts a compressed format into uncompressed video. There is a complex relationship between video quality, the amount of data used to represent the video (determined by the bit rate), the complexity of the encoding and decoding algorithms, susceptibility to data loss and errors, ease of editing, random access, and end-to-end delay (latency). The compression format typically conforms to a standard video compression specification, such as the High Efficiency Video Coding (HEVC) standard (also referred to as h.265 or MPEG-H part 2), the general video coding standard to be finalized, or other current and/or future video coding standards.
Sub-block based prediction was first introduced into the video coding standard by the High Efficiency Video Coding (HEVC) standard. With sub-block based prediction, a block, such as a Coding Unit (CU) or a Prediction Unit (PU), is divided into non-overlapping sub-blocks. Different sub-blocks may be assigned different motion information, such as reference indices or Motion Vectors (MVs), and Motion Compensation (MC) is performed separately for each sub-block. Fig. 1 shows an example of sub-block based prediction.
Embodiments of the disclosed techniques may be applied to existing video coding standards (e.g., HEVC, h.265) and future standards to improve runtime performance. Section headings are used in this document to enhance readability of the specification, and discussion or embodiments (and/or implementations) are not limited in any way to only the corresponding sections.
1. Example of Joint Exploration Model (JEM)
In some embodiments, reference software called Joint Exploration Model (JEM) is used to explore future video coding techniques. In JEM, sub-block based prediction is employed in several coding tools, such as affine prediction, Alternative Temporal Motion Vector Prediction (ATMVP), spatio-temporal motion vector prediction (STMVP), bi-directional optical flow (BIO), frame rate up-conversion (FRUC), Locally Adaptive Motion Vector Resolution (LAMVR), Overlapped Block Motion Compensation (OBMC), Local Illumination Compensation (LIC), and decoder-side motion vector refinement (DMVR).
1.1 example of affine prediction
In HEVC, only the translational motion model is applied to Motion Compensated Prediction (MCP). However, the camera and the object may have a variety of motions, such as zoom in/out, rotation, perspective motion, and/or other irregular motions. JEM, on the other hand, applies simplified affine transform motion compensated prediction. FIG. 2 shows a motion vector V from two control points 0 And V 1 An example of an affine motion field of block 200 is described. The Motion Vector Field (MVF) of block 200 may be described by:
as shown in fig. 2, (v) 0x ,v 0y ) Is to the leftMotion vector of upper corner control point, and (v) 1x ,v 1y ) Is the motion vector of the upper right hand corner control point. In order to simplify motion compensated prediction, sub-block based affine transform prediction may be applied. The subblock size M × N is derived as follows:
here, MvPre is the motion vector fractional accuracy (e.g., 1/16 in JEM). (v) of 2x ,v 2y ) Is the motion vector of the lower left control point, which is calculated according to equation (1). If desired, M and N can be adjusted downward to be divisors of w and h, respectively.
Fig. 3 shows an example of affine MVF for each sub-block of block 300. To derive the motion vector for each M × N sub-block, the motion vector for the center sample of each sub-block may be calculated according to equation (1) and rounded to the motion vector fractional accuracy (e.g., 1/16 in JEM). A motion compensated interpolation filter may then be applied to generate a prediction for each sub-block using the derived motion vectors. After MCP, the high accuracy motion vector of each sub-block is rounded and saved to the same accuracy as the normal motion vector.
In JEM, there are two affine motion patterns: AF _ INTER mode and AF _ MERGE mode. For CUs with width and height greater than 8, the AF _ INTER mode may be applied. An affine flag at the CU level is signaled in the bitstream to indicate whether AF _ INTER mode is used. In AF _ INTER mode, neighboring block construction is used with motion vector pair { (v) 0 ,v 1 )|v 0 ={v A ,v B ,v c },v 1 ={v D ,v E } of the candidate list.
Fig. 4 shows an example of Motion Vector Prediction (MVP) for a block 400 in the AF _ INTER mode. As shown in fig. 4, v is selected from the motion vectors of sub-block A, B or C 0 . The motion vectors from the neighboring blocks may be scaled according to the reference list. The reference Picture Order Count (POC) of the neighboring block, the reference POC of the current CU, and the POC of the current CU may also be based on a correlation between the POC of the current CU and the reference POC of the neighboring blockThe motion vectors are scaled. Selecting v from adjacent sub-blocks D and E 1 The method of (3) is similar. If the number of candidate lists is less than 2, the list is populated by pairs of motion vectors that are constructed by duplicating each Advanced Motion Vector Prediction (AMVP) candidate. When the candidate list is greater than 2, the candidates may first be filtered according to neighboring motion vectors (e.g., based on the similarity of two motion vectors in the candidates). In some embodiments, the first two candidates are retained. In some embodiments, a Rate Distortion (RD) cost check is used to determine which motion vector pair candidate to select as the Control Point Motion Vector Predictor (CPMVP) for the current CU. An index indicating the position of the CPMVP in the candidate list may be signaled in the bitstream. After determining the CPMVP of the current affine CU, affine motion estimation is applied and Control Point Motion Vectors (CPMVs) are found. The difference between CPMV and CPMVP is then signaled in the bitstream.
When a CU is applied in AF _ MERGE mode, it obtains the first block encoded in affine mode from the valid neighboring reconstructed blocks. Fig. 5A shows an example of the selection order of candidate blocks of the current CU 500. As shown in fig. 5A, the selection order may be from left (501), top (502), top right (503), bottom left (504), to top left (505) of the current CU 500. Fig. 5B shows another example of a candidate block of the current CU500 in AF _ MERGE mode. If the neighboring lower left block 501 is encoded in affine mode, as shown in fig. 5B, then the motion vectors v for the upper left, upper right and lower left corner of the CU containing sub-block 501 are derived 2 、v 3 And v 4 . Based on v 2 、v 3 And v 4 Calculating motion vector v of the top left corner on current CU500 0 . The motion vector v at the top right of the current CU can be calculated accordingly 1 。
Calculating the CPMV v of the current CU in accordance with the affine motion model in equation (1) 0 And v 1 Thereafter, the MVF of the current CU may be generated. To identify whether the current CU is encoded in AF _ MERGE mode, an affine flag may be signaled in the bitstream when there is at least one neighboring block encoded in affine mode.
In JEM, the non-Merge affine mode can only be used if the width and height of the current block are both greater than 8; the Merge affine mode can be used only when the area (i.e., width x height) of the current block is not less than 64.
2. Examples of existing methods for sub-block based implementations
In some prior implementations, the size of the sub-blocks (e.g., 4 × 4 in JEM) is designed primarily for the luma component. For example, in JEM, the size of the sub-block is for a block with 4: 2: 2 x 2 chroma components in 0 format, and for a chroma component having a 4: 2: 2 x 4 chrominance components in 2 format. The small size of the sub-blocks requires higher bandwidth requirements. FIG. 6 shows a JEM having 4: 2: example of sub-blocks of 16 x 16 blocks (8 x 8 for Cb/Cr) of different components of the 0 format.
In other prior implementations, in some sub-block based tools (e.g., affine prediction in JEM), the MV of each sub-block is calculated independently for each component using the affine model shown in equation (1), which may result in misalignment of the motion vector between the luma and chroma components.
In other existing implementations, in some sub-block based tools (e.g., affine prediction), the usage constraints are different for both the Merge mode and the non-Merge inter mode (also referred to as AMVP mode, or normal inter mode), which needs to be unified.
3. Exemplary method for subblock-based prediction in video coding
The sub-block based prediction method includes unifying constraints for Merge affine mode and non-Merge affine mode. The use of sub-block based prediction to improve video coding efficiency and enhance existing and future video coding standards is set forth in the examples described below for the various embodiments.
Example 1.The Merge affine mode and the non-Merge affine mode are allowed or not allowed under the same block size constraint.
(a) The block size constraint depends on the width and height compared to one or two thresholds. For example, if the width and height of the current block are both greater than M (e.g., M equals 8), or the width is greater than M0 and the height is greater than M1 (e.g., M0 equals 8 and M1 equals 4), then Merge affine mode and non-Merge affine mode are allowed; otherwise, Merge affine mode and non-Merge affine mode are not allowed. In another example, if the width and height of the current block are both greater than M (e.g., M equals 16), then both a Merge affine mode and a non-Merge affine mode are allowed; otherwise, Merge affine mode and non-Merge affine mode are not allowed.
(b) The block size constraint depends on the total number of samples within one block (i.e., the area width x height). In one example, if the area (i.e., width x height) of the current block is not less than N (e.g., N is equal to 64), then both the Merge affine mode and the non-Merge affine mode are allowed; otherwise, Merge affine mode and non-Merge affine mode are not allowed.
(c) For the Merge affine mode, it can be an explicit mode that signals a flag as in JEM, or it can be an implicit mode that does not signal a flag as in other embodiments. In the latter case, if the Merge affine mode is not allowed, the affine Merge candidates are not put into the unified Merge candidate list.
(d) For non-Merge affine modes, when affine is not allowed according to the above rules, the signaling of the indication of affine mode is skipped.
The above examples may be incorporated in the context of methods described below (e.g., methods 700 and 750), which methods 700 and 750 may be implemented at a video encoder and a video decoder, respectively.
Fig. 7A shows a flow diagram of an exemplary method for video processing. The method 700 includes determining a block size constraint for a video block in a video frame at operation 710. At operation 720, a determination is made as to whether a Merge affine mode and a non-Merge affine mode are allowed for the video block based on the block size constraint. At operation 730, a bit stream representation of the video block is generated based on the determination.
In some embodiments, generating the bitstream representation comprises encoding the video block using a Merge affine mode by including an indication of the Merge affine mode in the bitstream representation. In some embodiments, generating the bitstream representation comprises encoding the video block using the non-Merge affine mode by including an indication of the non-Merge affine mode in the bitstream representation. In some embodiments, generating the bitstream representation comprises encoding the video block using the Merge affine mode by omitting an indication of the Merge affine mode in the bitstream representation to implicitly indicate the Merge affine mode. In some embodiments, generating the bitstream representation comprises encoding the video block using the non-Merge affine mode by omitting the indication of the non-Merge affine mode in the bitstream representation to implicitly indicate the non-Merge affine mode.
In some embodiments, generating the bitstream representation comprises generating the bitstream representation from the video block such that the bitstream omits an explicit indication of the Merge affine mode due to determining that the block size constraint does not allow the Merge affine mode. In some embodiments, generating the bitstream representation comprises generating the bitstream representation from the video block such that the bitstream omits an explicit indication of the non-Merge affine mode due to determining that the block size constraint does not allow the non-Merge affine mode to be passed. In some embodiments, the block size constraints for the video blocks include a height of the video blocks and a width of the video blocks, both of which are greater than a common threshold. In some embodiments, the common threshold is eight, and the Merge affine mode and the non-Merge affine mode are allowed in response to the height of the video block and the width of the video block being greater than the common threshold. In some embodiments, the common threshold is 16, and the Merge affine mode and the non-Merge affine mode are allowed in response to the height of the video block and the width of the video block being greater than the common threshold.
In some embodiments, the block size constraints for the video blocks include a height of the video block being greater than a first threshold and a width of the video block being greater than a second threshold, and the first threshold and the second threshold being different. In some embodiments, the block size constraints for the video blocks include a height of the video blocks and a width of the video blocks, and a product of the height and the width is greater than a threshold.
Fig. 7B illustrates another flow diagram of an exemplary method for motion compensation. The method 750 includes determining block size constraints for video blocks in a video frame at operation 760. At operation 770, a determination is made as to whether the video block allows for the Merge affine mode and the non-Merge affine mode based on the block size constraint. In operation 780, a video block is generated from the bit stream representation of the video block based on the determining.
In some embodiments, the bitstream representation includes an indication of a Merge affine mode, and the video block is generated using the Merge affine mode by parsing the bitstream representation according to the indication. In some embodiments, the bitstream representation includes an indication of a non-Merge affine mode, and the video chunk is generated using the non-Merge affine mode by parsing the bitstream representation according to the indication. In some embodiments, the bitstream representation omits an indication of the Merge affine mode, and wherein the video block is generated using the Merge affine mode by parsing the bitstream representation without the indication of the Merge affine mode. In some embodiments, the bitstream representation omits the indication of the non-Merge affine mode and the video block is generated using the non-Merge affine mode by parsing the bitstream representation without the indication of the non-Merge affine mode.
In some embodiments, the block size constraints for the video blocks include heights of the video blocks and widths of the video blocks, the heights and widths being greater than a common threshold. In some embodiments, the block size constraints for the video blocks include that the height of the video blocks is greater than a first threshold and the width of the video blocks is greater than a second threshold, and the first and second thresholds are different. In some embodiments, the block size constraints for the video blocks include a height of the video blocks and a width of the video blocks, and a product of the height and the width is greater than a threshold.
In a representative aspect, the methods 700 and/or 750 described above are embodied in processor executable code and stored in a computer readable program medium. In yet another representative aspect, an apparatus configured or operable to perform the above-described method is disclosed. The device may include a processor programmed to implement the methods 700 and/or 750 described above. In yet another representative aspect, a video encoder device may implement method 700. In yet another representative aspect, a video decoder device may implement method 750.
4. Example embodiments of the disclosed technology
Fig. 8 is a block diagram illustrating an example of an architecture of a computer system or other control device that may be used to implement various portions of the presently disclosed technology, including (but not limited to) methods 700 and/or 750. In fig. 8, a computer system 800 includes one or more processors 805 and memory 810 connected via an interconnect 825. Interconnect 825 may represent any one or more separate physical buses, point-to-point connections, or both, connected through appropriate bridges, adapters, or controllers. Thus, interconnect 825 may comprise, for example, a system bus, a Peripheral Component Interconnect (PCI) bus, a HyperTransport or Industry Standard Architecture (ISA) bus, a Small Computer System Interface (SCSI) bus, a Universal Serial Bus (USB), an IIC (I2C) bus, or an Institute of Electrical and Electronics Engineers (IEEE) standard 674 bus, also sometimes referred to as a "Firewire".
The processor(s) 805 may include a Central Processing Unit (CPU) to control overall operation of, for example, a host computer. In certain embodiments, the processor(s) 805 achieve this by executing software or firmware stored in memory 810. The processor(s) 805 may be or include one or more programmable general-purpose or special-purpose microprocessors, Digital Signal Processors (DSPs), programmable controllers, Application Specific Integrated Circuits (ASICs), Programmable Logic Devices (PLDs), or the like, or a combination of such devices.
The memory 810 may be or include the main memory of a computer system. Memory 810 represents any suitable form of Random Access Memory (RAM), Read Only Memory (ROM), flash memory, etc., or combination of such devices. In use, the memory 810 may contain, among other things, a set of machine instructions that, when executed by the processor 805, cause the processor 805 to perform operations to implement embodiments of the presently disclosed technology.
An (optional) network adapter 815 is also connected to the processor(s) 805 via the interconnect 825. The network adapter 815 provides the computer system 800 with the ability to communicate with remote devices, such as storage clients and/or other storage servers, and the network adapter 815 may be, for example, an ethernet adapter or a fibre channel adapter.
Fig. 9 illustrates a block diagram of an example embodiment of a mobile device 900, which mobile device 900 may be used to implement various portions of the presently disclosed technology, including (but not limited to) methods 700 and/or 750. The mobile device 900 may be a laptop computer, a smart phone, a tablet computer, a camcorder, or other device capable of processing video. The mobile device 900 includes a processor or controller 901 for processing data and memory 902 in communication with the processor 901 for storing and/or buffering data. For example, the processor 901 may include a Central Processing Unit (CPU) or a microcontroller unit (MCU). In some implementations, the processor 901 may include a Field Programmable Gate Array (FPGA). In some implementations, the mobile device 900 includes or communicates with a Graphics Processing Unit (GPU), a Video Processing Unit (VPU), and/or a wireless communication unit for various visual and/or communication data processing functions of a smartphone device. For example, the memory 902 may include and store processor-executable code that, when executed by the processor 901, configures the mobile device 900 to perform various operations, such as receiving information, commands, and/or data, processing information and data, and transmitting or providing processed information/data to another device, such as an actuator or external display.
To support various functions of the mobile device 900, the memory 902 can store information and data such as instructions, software, values, images, and other data that are processed or referenced by the processor 901. For example, various types of Random Access Memory (RAM) devices, Read Only Memory (ROM) devices, flash memory devices, and other suitable storage media may be used to implement the storage functionality of memory 902. In some implementations, the mobile device 900 includes an input/output (I/O) unit 903 to connect the processor 901 and/or the memory 902 to other modules, units, or devices. For example, the I/O unit 903 may be connected with the processor 901 and the memory 902 to utilize various types of wireless interfaces compatible with general data communication standards, for example, between one or more computers in the cloud and a user device. In some implementations, the mobile device 900 can connect with other devices using a wired connection via the I/O unit 903. The mobile device 900 may also be connected to other external interfaces, such as data storage and/or a visual or audio display device 904, to retrieve and transfer data and information which may be processed by the processor, stored in memory, or presented on an output unit of the display device 904 or an external device. For example, display device 904 may display a video frame that includes blocks (CU, PU, or TU) that apply intra block copying based on whether the blocks are encoded using a motion compensation algorithm and in accordance with the disclosed techniques.
In some embodiments, a video encoder device or a decoder device may implement the method of sub-block based prediction as described herein for video encoding or decoding. Various features of the method may be similar to methods 700 or 750 described above.
In some embodiments, the video encoding and/or decoding methods may be implemented using decoding devices implemented on the hardware platforms described with respect to fig. 8 and 9.
From the foregoing it will be appreciated that specific embodiments of the presently disclosed technology have been described herein for purposes of illustration, but that various modifications may be made without deviating from the scope of the invention. Accordingly, the presently disclosed technology is not limited, except as by the appended claims.
Embodiments of the subject matter and the functional operations described in this patent document can be implemented in various systems, digital electronic circuitry, or computer software, firmware, or hardware, including the structures disclosed in this specification and their structural equivalents, or combinations of one or more of them. Embodiments of the subject matter described in this specification can be implemented as one or more computer program products, i.e., one or more modules of computer program instructions encoded on a tangible and non-transitory computer readable medium for execution by, or to control the operation of, data processing apparatus. The computer readable medium can be a machine-readable storage device, a machine-readable storage substrate, a memory device, a composition of matter effecting a machine-readable propagated signal, or a combination of one or more of them. The term "data processing unit" or "data processing apparatus" includes all apparatus, devices, and machines for processing data, including by way of example a programmable processor, a computer, or multiple processors or computers. The apparatus can include, in addition to hardware, code that creates an execution environment for the computer program in question, e.g., code that constitutes processor firmware, a protocol stack, a database management system, an operating system, or a combination of one or more of them.
A computer program (also known as a program, software application, script, or code) can be written in any form of programming language, including compiled or interpreted languages, and it can be deployed in any form, including as a stand-alone program or as a module, component, subroutine, or other unit suitable for use in a computing environment. The computer program does not necessarily correspond to a file in a file system. A program can be stored in a portion of a file that holds other programs or data (e.g., one or more scripts stored in a markup language document), in a single file dedicated to the program in question, or in multiple coordinated files (e.g., files that store one or more modules, sub programs, or portions of code). A computer program can be deployed to be executed on one computer or on multiple computers that are located at one site or distributed across multiple sites and interconnected by a communication network.
The processes and logic flows described in this specification can be performed by one or more programmable processors executing one or more computer programs to perform functions by operating on input data and generating output. The processes and logic flows can also be performed by, and apparatus can also be implemented as, special purpose logic circuitry, e.g., an FPGA (field programmable gate array) or an ASIC (application-specific integrated circuit).
Processors suitable for the execution of a computer program include, by way of example, both general and special purpose microprocessors, and any one or more processors of any kind of digital computer. Generally, a processor will receive instructions and data from a read-only memory or a random access memory or both. The essential elements of a computer are a processor for executing instructions and one or more memory devices for storing instructions and data. Generally, a computer will also include, or be operatively coupled to receive data from or transfer data to, or both, one or more mass storage devices for storing data, e.g., magnetic, magneto-optical disks, or optical disks. However, a computer need not have such devices. Computer-readable media suitable for storing computer program instructions and data include all forms of non-volatile memory, media and memory devices, including by way of example semiconductor memory devices, e.g., EPROM, EEPROM, and flash memory devices. The processor and the memory can be supplemented by, or incorporated in, special purpose logic circuitry.
The specification, together with the drawings, should be considered exemplary only, with the examples being meant as examples. As used herein, the singular forms "a", "an" and "the" are intended to include the plural forms as well, unless the context clearly indicates otherwise. In addition, the use of "or" is intended to include "and/or" unless the context clearly indicates otherwise.
While this patent document contains many specifics, these should not be construed as limitations on the scope of any invention or of what may be claimed, but rather as descriptions of features specific to particular embodiments of particular inventions. Certain features that are described in this patent document in the context of separate embodiments can also be implemented in combination in a single embodiment. Conversely, various features that are described in the context of a single embodiment can also be implemented in multiple embodiments separately or in any suitable subcombination. Furthermore, although features may be described above as acting in certain combinations and even initially claimed as such, one or more features from a claimed combination can in some cases be excised from the combination, and the claimed combination may be directed to a subcombination or variation of a subcombination.
Similarly, while operations are depicted in the drawings in a particular order, this should not be understood as requiring that such operations be performed in the particular order shown or in sequential order, or that all illustrated operations be performed, to achieve desirable results. Moreover, the separation of various system components in the embodiments described in this patent document should not be understood as requiring such separation in all embodiments.
Only a few embodiments and examples are described and other embodiments, enhancements and variations can be made based on what is described and illustrated in this patent document.
Claims (20)
1. A method for video processing, comprising:
determining whether a Merge affine mode and a non-Merge affine mode are allowed for a video block in a video frame based on a uniform block size constraint; and
generating the video block from a bit stream of the video block based on making the determination, wherein,
determining that the non-Merge affine mode and the Merge affine mode are not allowed in response to the height of the video block and the width of the video block being less than 16, thus omitting an indication of the non-Merge affine mode and the Merge affine mode in the bitstream,
generating the video block by parsing the bitstream without the indication of the non-Merge affine mode and the Merge affine mode and without affine mode in the bitstream if the indication of the non-Merge affine mode and the Merge affine mode is omitted.
2. The method of claim 1, wherein the first and second light sources are selected from the group consisting of,
wherein, in case it is determined that the Merge affine mode is allowed, including an indication of the Merge affine mode in the bitstream, by parsing the bitstream according to the indication and generating the video block using the Merge affine mode.
3. The method of claim 1, wherein the first and second light sources are selected from the group consisting of,
wherein, in a case where it is determined that the non-Merge affine mode is allowed, in a case where an indication of the non-Merge affine mode is included in the bitstream, the video block is generated by parsing the bitstream according to the indication and using the non-Merge affine mode.
4. The method of claim 1, wherein the block size constraints of the video block comprise a height of the video block and a width of the video block, the height and the width being greater than a common threshold.
5. The method of claim 4, wherein the first and second light sources are selected from the group consisting of,
wherein the common threshold is 16, and
wherein the Merge affine mode and the non-Merge affine mode are allowed in response to the height of the video block and the width of the video block being greater than the common threshold.
6. The method of claim 1, wherein the first and second light sources are selected from the group consisting of,
wherein the block size constraint for the video block comprises a height of the video block being greater than a first threshold and a width of the video block being greater than a second threshold, and
wherein the first threshold and the second threshold are different.
7. The method of claim 1, wherein the first and second light sources are selected from the group consisting of,
wherein the block size constraints for the video block include a height of the video block and a width of the video block, and
wherein a product of the height and the width is greater than a threshold.
8. A method for video processing, comprising:
determining whether a Merge affine mode and a non-Merge affine mode are allowed for a video block in a video frame based on a uniform block size constraint; and
generating a bitstream for the video block based on the determining, wherein,
generating the bitstream from the video block such that the non-Merge affine mode and the Merge affine mode are determined not to be allowed in response to the height of the video block and the width of the video block being less than 16, thereby omitting an indication of the non-Merge affine mode and the Merge affine mode in the bitstream,
encoding the video block by not using an affine mode without the indication of the non-Merge affine mode and the Merge affine mode if the indication of the non-Merge affine mode and the Merge affine mode is omitted in the bitstream.
9. The method of claim 8, wherein, if it is determined that the Merge affine mode is allowed, including an indication of the Merge affine mode in the bitstream, the video block is encoded using the Merge affine mode.
10. The method of claim 8, wherein, if the non-Merge affine mode is determined to be allowed, including an indication of the non-Merge affine mode in the bitstream, the video block is encoded using the non-Merge affine mode.
11. The method of claim 8, wherein the block size constraints of the video block comprise a height of the video block and a width of the video block, the height and the width both being greater than a common threshold.
12. The method of claim 11, wherein the first and second light sources are selected from the group consisting of,
wherein the common threshold is 16, and
wherein the Merge affine mode and the non-Merge affine mode are allowed in response to the height of the video block and the width of the video block being greater than the common threshold.
13. The method of claim 8, wherein the first and second light sources are selected from the group consisting of,
wherein the block size constraint for the video block comprises a height of the video block being greater than a first threshold and a width of the video block being greater than a second threshold, and
wherein the first threshold and the second threshold are different.
14. The method of claim 8, wherein the first and second light sources are selected from the group consisting of,
wherein the block size constraints for the video block include a height of the video block and a width of the video block, and
wherein a product of the height and the width is greater than a threshold.
15. A video decoding apparatus, comprising: a processor and a non-transitory memory having instructions thereon, wherein the instructions, when executed by the processor, cause the processor to implement the method of any one of claims 1 to 7.
16. A video encoding device, comprising: a processor and a non-transitory memory having instructions thereon, wherein the instructions, when executed by the processor, cause the processor to implement the method of any one of claims 8-14.
17. An apparatus for video processing, comprising:
the determining module is used for determining whether the Merge affine mode and the non-Merge affine mode are allowed for the video block in the video frame based on the unified block size constraint; and
a generation module to generate the video block from a bitstream of the video block based on the determining, wherein,
in response to the height of the video block and the width of the video block being less than 16, determining that the non-Merge affine mode and the Merge affine mode are not allowed, thus omitting an indication of the non-Merge affine mode and the Merge affine mode in the bitstream,
generating the video block by parsing the bitstream without the indication of the non-Merge affine mode and the Merge affine mode and without affine mode in the bitstream if the indication of the non-Merge affine mode and the Merge affine mode is omitted.
18. An apparatus for video processing, comprising:
the determining module is used for determining whether the Merge affine mode and the non-Merge affine mode are allowed for the video blocks in the video frame based on the uniform block size constraint; and
a generation module to generate a bitstream for the video block based on the determining, wherein,
generating the bitstream from the video block such that the non-Merge affine mode and the Merge affine mode are determined not to be allowed in response to a height of the video block and a width of the video block being less than 16, thereby omitting an indication of the non-Merge affine mode and the Merge affine mode in the bitstream,
encoding the video block by not using an affine mode without the indication of the non-Merge affine mode and the Merge affine mode if the indication of the non-Merge affine mode and the Merge affine mode is omitted in the bitstream.
19. A non-transitory computer-readable storage medium for storing program code, wherein the program code, when executed by a computer, the computer implements the method of any of claims 1-14.
20. A method for storing a video bitstream, comprising:
determining whether a Merge affine mode and a non-Merge affine mode are allowed for a video block in a video frame based on a uniform block size constraint;
based on the determining, generating a bitstream for the video block; and
storing the bitstream in a non-transitory computer-readable storage medium, wherein,
generating the bitstream from the video block such that the non-Merge affine mode and the Merge affine mode are determined not to be allowed in response to a height of the video block and a width of the video block being less than 16, thereby omitting an indication of the non-Merge affine mode and the Merge affine mode in the bitstream,
encoding the video block by not using an affine mode without the indication of the non-Merge affine mode and the Merge affine mode if the indication of the non-Merge affine mode and the Merge affine mode is omitted in the bitstream.
Applications Claiming Priority (2)
Application Number | Priority Date | Filing Date | Title |
---|---|---|---|
CNPCT/CN2018/092118 | 2018-06-21 | ||
CN2018092118 | 2018-06-21 |
Publications (2)
Publication Number | Publication Date |
---|---|
CN110636298A CN110636298A (en) | 2019-12-31 |
CN110636298B true CN110636298B (en) | 2022-09-13 |
Family
ID=67874478
Family Applications (1)
Application Number | Title | Priority Date | Filing Date |
---|---|---|---|
CN201910544642.0A Active CN110636298B (en) | 2018-06-21 | 2019-06-21 | Unified constraints for Merge affine mode and non-Merge affine mode |
Country Status (4)
Country | Link |
---|---|
US (1) | US11197003B2 (en) |
CN (1) | CN110636298B (en) |
TW (1) | TWI739120B (en) |
WO (1) | WO2019244117A1 (en) |
Families Citing this family (21)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
WO2020065520A2 (en) | 2018-09-24 | 2020-04-02 | Beijing Bytedance Network Technology Co., Ltd. | Extended merge prediction |
WO2019234598A1 (en) | 2018-06-05 | 2019-12-12 | Beijing Bytedance Network Technology Co., Ltd. | Interaction between ibc and stmvp |
GB2589223B (en) | 2018-06-21 | 2023-01-25 | Beijing Bytedance Network Tech Co Ltd | Component-dependent sub-block dividing |
CN111010571B (en) | 2018-10-08 | 2023-05-16 | 北京字节跳动网络技术有限公司 | Generation and use of combined affine Merge candidates |
WO2020084476A1 (en) | 2018-10-22 | 2020-04-30 | Beijing Bytedance Network Technology Co., Ltd. | Sub-block based prediction |
KR20210089155A (en) | 2018-11-10 | 2021-07-15 | 베이징 바이트댄스 네트워크 테크놀로지 컴퍼니, 리미티드 | Rounding in the fairwise average candidate calculation |
WO2020098655A1 (en) | 2018-11-12 | 2020-05-22 | Beijing Bytedance Network Technology Co., Ltd. | Motion vector storage for inter prediction |
CN117319644A (en) | 2018-11-20 | 2023-12-29 | 北京字节跳动网络技术有限公司 | Partial position based difference calculation |
WO2020103870A1 (en) | 2018-11-20 | 2020-05-28 | Beijing Bytedance Network Technology Co., Ltd. | Inter prediction with refinement in video processing |
WO2020103940A1 (en) | 2018-11-22 | 2020-05-28 | Beijing Bytedance Network Technology Co., Ltd. | Coordination method for sub-block based inter prediction |
JP2022521554A (en) | 2019-03-06 | 2022-04-08 | 北京字節跳動網絡技術有限公司 | Use of converted one-sided prediction candidates |
CN113647099B (en) | 2019-04-02 | 2022-10-04 | 北京字节跳动网络技术有限公司 | Decoder-side motion vector derivation |
CN113906753B (en) | 2019-04-24 | 2023-12-01 | 字节跳动有限公司 | Constraint of quantized residual differential pulse codec modulation representation for codec video |
KR20220002292A (en) * | 2019-05-01 | 2022-01-06 | 바이트댄스 아이엔씨 | Intra-coded video using quantized residual differential pulse code modulation coding |
KR20220002917A (en) | 2019-05-02 | 2022-01-07 | 바이트댄스 아이엔씨 | Coding mode based on coding tree structure type |
CN113906759A (en) | 2019-05-21 | 2022-01-07 | 北京字节跳动网络技术有限公司 | Syntax-based motion candidate derivation in sub-block Merge mode |
KR20220043109A (en) | 2019-08-13 | 2022-04-05 | 베이징 바이트댄스 네트워크 테크놀로지 컴퍼니, 리미티드 | Motion precision of sub-block-based inter prediction |
US20220337814A1 (en) * | 2019-09-19 | 2022-10-20 | Lg Electronics Inc. | Image encoding/decoding method and device using reference sample filtering, and method for transmitting bitstream |
CN114424553A (en) | 2019-09-22 | 2022-04-29 | 北京字节跳动网络技术有限公司 | Inter-frame prediction scaling method based on sub-blocks |
MX2022004409A (en) | 2019-10-18 | 2022-05-18 | Beijing Bytedance Network Tech Co Ltd | Syntax constraints in parameter set signaling of subpictures. |
CN112788345B (en) * | 2019-11-11 | 2023-10-24 | 腾讯美国有限责任公司 | Video data decoding method, device, computer equipment and storage medium |
Citations (3)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
CN106537915A (en) * | 2014-07-18 | 2017-03-22 | 联发科技(新加坡)私人有限公司 | Method of motion vector derivation for video coding |
CN106559669A (en) * | 2015-09-29 | 2017-04-05 | 华为技术有限公司 | The method and device of image prediction |
WO2017118411A1 (en) * | 2016-01-07 | 2017-07-13 | Mediatek Inc. | Method and apparatus for affine inter prediction for video coding system |
Family Cites Families (161)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
DE60024389T2 (en) | 1999-04-26 | 2006-08-03 | Koninklijke Philips Electronics N.V. | SUBPIXEL ACCURATE MOTION VECTOR ESTIMATION AND MOTION COMPENSATED INTERPOLATION |
CN1311409C (en) | 2002-07-31 | 2007-04-18 | 皇家飞利浦电子股份有限公司 | System and method for segmenting |
CN1777283A (en) | 2004-12-31 | 2006-05-24 | 上海广电(集团)有限公司 | Microblock based video signal coding/decoding method |
US8954943B2 (en) | 2006-01-26 | 2015-02-10 | International Business Machines Corporation | Analyze and reduce number of data reordering operations in SIMD code |
US8184715B1 (en) | 2007-08-09 | 2012-05-22 | Elemental Technologies, Inc. | Method for efficiently executing video encoding operations on stream processor architectures |
US20110002386A1 (en) | 2009-07-06 | 2011-01-06 | Mediatek Singapore Pte. Ltd. | Video encoder and method for performing intra-prediction and video data compression |
JP5234368B2 (en) | 2009-09-30 | 2013-07-10 | ソニー株式会社 | Image processing apparatus and method |
EP2532159A1 (en) | 2010-02-05 | 2012-12-12 | Telefonaktiebolaget L M Ericsson (PUBL) | Selecting predicted motion vector candidates |
US20120287999A1 (en) | 2011-05-11 | 2012-11-15 | Microsoft Corporation | Syntax element prediction in error correction |
US9866859B2 (en) | 2011-06-14 | 2018-01-09 | Texas Instruments Incorporated | Inter-prediction candidate index coding independent of inter-prediction candidate list construction in video coding |
GB201113527D0 (en) | 2011-08-04 | 2011-09-21 | Imagination Tech Ltd | External vectors in a motion estimation system |
KR20140057373A (en) | 2011-08-30 | 2014-05-12 | 노키아 코포레이션 | An apparatus, a method and a computer program for video coding and decoding |
MX353235B (en) | 2011-09-29 | 2018-01-08 | Sharp Kk Star | Image decoding device, image decoding method, and image encoding device. |
JP5895469B2 (en) | 2011-11-18 | 2016-03-30 | 富士通株式会社 | Video encoding device and video decoding device |
KR20130058524A (en) | 2011-11-25 | 2013-06-04 | 오수미 | Method for generating chroma intra prediction block |
US9451252B2 (en) | 2012-01-14 | 2016-09-20 | Qualcomm Incorporated | Coding parameter sets and NAL unit headers for video coding |
CN110830799B (en) | 2012-01-18 | 2023-09-15 | 韩国电子通信研究院 | Video decoding device, video encoding device and method for transmitting bit stream |
CN104221376B (en) | 2012-04-12 | 2017-08-29 | 寰发股份有限公司 | The method and apparatus that video data is handled in video coding system |
WO2014007521A1 (en) | 2012-07-02 | 2014-01-09 | 삼성전자 주식회사 | Method and apparatus for predicting motion vector for coding video or decoding video |
EP3588958A1 (en) | 2012-08-29 | 2020-01-01 | Vid Scale, Inc. | Method and apparatus of motion vector prediction for scalable video coding |
US9491461B2 (en) | 2012-09-27 | 2016-11-08 | Qualcomm Incorporated | Scalable extensions to HEVC and temporal motion vector prediction |
KR20150038249A (en) | 2012-09-28 | 2015-04-08 | 인텔 코포레이션 | Inter-layer pixel sample prediction |
KR102367210B1 (en) | 2012-10-01 | 2022-02-24 | 지이 비디오 컴프레션, 엘엘씨 | Scalable video coding using inter-layer prediction of spatial intra prediction parameters |
US9615089B2 (en) | 2012-12-26 | 2017-04-04 | Samsung Electronics Co., Ltd. | Method of encoding and decoding multiview video sequence based on adaptive compensation of local illumination mismatch in inter-frame prediction |
US9294777B2 (en) | 2012-12-30 | 2016-03-22 | Qualcomm Incorporated | Progressive refinement with temporal scalability support in video coding |
US9674542B2 (en) | 2013-01-02 | 2017-06-06 | Qualcomm Incorporated | Motion vector prediction for video coding |
US20140254678A1 (en) | 2013-03-11 | 2014-09-11 | Aleksandar Beric | Motion estimation using hierarchical phase plane correlation and block matching |
US9521425B2 (en) | 2013-03-19 | 2016-12-13 | Qualcomm Incorporated | Disparity vector derivation in 3D video coding for skip and direct modes |
WO2014166116A1 (en) | 2013-04-12 | 2014-10-16 | Mediatek Inc. | Direct simplified depth coding |
US10045014B2 (en) | 2013-07-15 | 2018-08-07 | Mediatek Singapore Pte. Ltd. | Method of disparity derived depth coding in 3D video coding |
US9628795B2 (en) | 2013-07-17 | 2017-04-18 | Qualcomm Incorporated | Block identification using disparity vector in video coding |
WO2015006967A1 (en) | 2013-07-19 | 2015-01-22 | Mediatek Singapore Pte. Ltd. | Simplified view synthesis prediction for 3d video coding |
CN104769947B (en) | 2013-07-26 | 2019-02-26 | 北京大学深圳研究生院 | A kind of more hypothesis motion compensation encoding methods based on P frame |
WO2015010317A1 (en) | 2013-07-26 | 2015-01-29 | 北京大学深圳研究生院 | P frame-based multi-hypothesis motion compensation method |
AU2013228045A1 (en) | 2013-09-13 | 2015-04-02 | Canon Kabushiki Kaisha | Method, apparatus and system for encoding and decoding video data |
US9667996B2 (en) | 2013-09-26 | 2017-05-30 | Qualcomm Incorporated | Sub-prediction unit (PU) based temporal motion vector prediction in HEVC and sub-PU design in 3D-HEVC |
US9762927B2 (en) | 2013-09-26 | 2017-09-12 | Qualcomm Incorporated | Sub-prediction unit (PU) based temporal motion vector prediction in HEVC and sub-PU design in 3D-HEVC |
WO2015093449A1 (en) | 2013-12-19 | 2015-06-25 | シャープ株式会社 | Merge-candidate derivation device, image decoding device, and image encoding device |
TWI536811B (en) | 2013-12-27 | 2016-06-01 | 財團法人工業技術研究院 | Method and system for image processing, decoding method, encoder and decoder |
EP4096221A1 (en) | 2014-01-03 | 2022-11-30 | Microsoft Technology Licensing, LLC | Block vector prediction in video and image coding/decoding |
WO2015109598A1 (en) | 2014-01-27 | 2015-07-30 | Mediatek Singapore Pte. Ltd. | Methods for motion parameter hole filling |
CN108965888B (en) | 2014-03-19 | 2021-05-04 | 株式会社Kt | Method of generating merge candidate list for multi-view video signal and decoding apparatus |
WO2015169200A1 (en) | 2014-05-06 | 2015-11-12 | Mediatek Singapore Pte. Ltd. | Method of block vector prediction for intra block copy mode coding |
US10327002B2 (en) | 2014-06-19 | 2019-06-18 | Qualcomm Incorporated | Systems and methods for intra-block copy |
KR102413529B1 (en) | 2014-06-19 | 2022-06-24 | 마이크로소프트 테크놀로지 라이센싱, 엘엘씨 | Unified intra block copy and inter prediction modes |
US20150373350A1 (en) | 2014-06-20 | 2015-12-24 | Qualcomm Incorporated | Temporal motion vector prediction (tmvp) indication in multi-layer codecs |
EP3175618A1 (en) | 2014-09-11 | 2017-06-07 | Euclid Discoveries, LLC | Perceptual optimization for model-based video encoding |
CN107005708A (en) | 2014-09-26 | 2017-08-01 | Vid拓展公司 | Decoding is replicated in the block of use time block vector forecasting |
US9918105B2 (en) | 2014-10-07 | 2018-03-13 | Qualcomm Incorporated | Intra BC and inter unification |
CN111741312B (en) | 2014-10-31 | 2024-03-19 | 三星电子株式会社 | Method and apparatus for encoding/decoding motion vector |
EP3202143B8 (en) | 2014-11-18 | 2019-09-25 | MediaTek Inc. | Method of bi-prediction video coding based on motion vectors from uni-prediction and merge candidate |
WO2016090568A1 (en) | 2014-12-10 | 2016-06-16 | Mediatek Singapore Pte. Ltd. | Binary tree block partitioning structure |
US11477477B2 (en) | 2015-01-26 | 2022-10-18 | Qualcomm Incorporated | Sub-prediction unit based advanced temporal motion vector prediction |
CN107431817B (en) | 2015-01-29 | 2020-03-24 | Vid拓展公司 | Method and apparatus for palette coding |
JP2018050091A (en) | 2015-02-02 | 2018-03-29 | シャープ株式会社 | Image decoder, image encoder, and prediction vector conducting device |
WO2016138513A1 (en) | 2015-02-27 | 2016-09-01 | Arris Enterprises, Inc. | Modification of unification of intra block copy and inter signaling related syntax and semantics |
WO2016165069A1 (en) | 2015-04-14 | 2016-10-20 | Mediatek Singapore Pte. Ltd. | Advanced temporal motion vector prediction in video coding |
CA2983881C (en) | 2015-04-29 | 2019-11-19 | Hfi Innovation Inc. | Method and apparatus for intra block copy reference list construction |
US20160337662A1 (en) | 2015-05-11 | 2016-11-17 | Qualcomm Incorporated | Storage and signaling resolutions of motion vectors |
GB2539213A (en) | 2015-06-08 | 2016-12-14 | Canon Kk | Schemes for handling an AMVP flag when implementing intra block copy coding mode |
KR102216947B1 (en) | 2015-06-08 | 2021-02-18 | 브이아이디 스케일, 인크. | Intra block copy mode for coding screen content |
US10148977B2 (en) | 2015-06-16 | 2018-12-04 | Futurewei Technologies, Inc. | Advanced coding techniques for high efficiency video coding (HEVC) screen content coding (SCC) extensions |
KR102264767B1 (en) | 2015-07-27 | 2021-06-14 | 미디어텍 인크. | Method of system for video coding using intra block copy mode |
KR102531222B1 (en) | 2015-08-25 | 2023-05-10 | 인터디지털 매디슨 페턴트 홀딩스 에스에이에스 | Inverse tone mapping based on luminance zones |
CN107925775A (en) | 2015-09-02 | 2018-04-17 | 联发科技股份有限公司 | The motion compensation process and device of coding and decoding video based on bi-directional predicted optic flow technique |
WO2017041271A1 (en) | 2015-09-10 | 2017-03-16 | Mediatek Singapore Pte. Ltd. | Efficient context modeling for coding a block of data |
WO2017076221A1 (en) | 2015-11-05 | 2017-05-11 | Mediatek Inc. | Method and apparatus of inter prediction using average motion vector for video coding |
CN105306944B (en) | 2015-11-30 | 2018-07-06 | 哈尔滨工业大学 | Chromatic component Forecasting Methodology in hybrid video coding standard |
US9955186B2 (en) | 2016-01-11 | 2018-04-24 | Qualcomm Incorporated | Block size decision for video coding |
WO2017130696A1 (en) | 2016-01-29 | 2017-08-03 | シャープ株式会社 | Prediction image generation device, moving image decoding device, and moving image encoding device |
US11109061B2 (en) | 2016-02-05 | 2021-08-31 | Mediatek Inc. | Method and apparatus of motion compensation based on bi-directional optical flow techniques for video coding |
US10368083B2 (en) | 2016-02-15 | 2019-07-30 | Qualcomm Incorporated | Picture order count based motion vector pruning |
JP6379186B2 (en) | 2016-02-17 | 2018-08-22 | テレフオンアクチーボラゲット エルエム エリクソン(パブル) | Method and apparatus for encoding and decoding video pictures |
EP3417618A4 (en) | 2016-02-17 | 2019-07-24 | Telefonaktiebolaget LM Ericsson (publ) | Methods and devices for encoding and decoding video pictures |
WO2017156669A1 (en) | 2016-03-14 | 2017-09-21 | Mediatek Singapore Pte. Ltd. | Methods for motion vector storage in video coding |
CN114466193A (en) | 2016-03-16 | 2022-05-10 | 联发科技股份有限公司 | Method and apparatus for pattern-based motion vector derivation for video coding |
US11223852B2 (en) | 2016-03-21 | 2022-01-11 | Qualcomm Incorporated | Coding video data using a two-level multi-type-tree framework |
US10567759B2 (en) | 2016-03-21 | 2020-02-18 | Qualcomm Incorporated | Using luma information for chroma prediction with separate luma-chroma framework in video coding |
WO2017164297A1 (en) | 2016-03-25 | 2017-09-28 | パナソニックIpマネジメント株式会社 | Method and device for encoding of video using signal dependent-type adaptive quantization and decoding |
WO2017171107A1 (en) | 2016-03-28 | 2017-10-05 | 엘지전자(주) | Inter-prediction mode based image processing method, and apparatus therefor |
CN116546206A (en) | 2016-04-08 | 2023-08-04 | 韩国电子通信研究院 | Method and apparatus for deriving motion prediction information |
WO2017188509A1 (en) | 2016-04-28 | 2017-11-02 | 엘지전자(주) | Inter prediction mode-based image processing method and apparatus therefor |
KR20190015216A (en) | 2016-05-05 | 2019-02-13 | 브이아이디 스케일, 인크. | Intra-directional representation based on control points for intra coding |
EP3457696A4 (en) | 2016-05-13 | 2019-12-18 | Sharp Kabushiki Kaisha | Predicted image generation device, video decoding device and video encoding device |
US10560718B2 (en) | 2016-05-13 | 2020-02-11 | Qualcomm Incorporated | Merge candidates for motion vector prediction for video coding |
US10560712B2 (en) | 2016-05-16 | 2020-02-11 | Qualcomm Incorporated | Affine motion prediction for video coding |
US9948930B2 (en) | 2016-05-17 | 2018-04-17 | Arris Enterprises Llc | Template matching for JVET intra prediction |
US20170339405A1 (en) | 2016-05-20 | 2017-11-23 | Arris Enterprises Llc | System and method for intra coding |
CA3025490A1 (en) | 2016-05-28 | 2017-12-07 | Mediatek Inc. | Method and apparatus of current picture referencing for video coding using affine motion compensation |
US10326986B2 (en) | 2016-08-15 | 2019-06-18 | Qualcomm Incorporated | Intra video coding using a decoupled tree structure |
US10368107B2 (en) | 2016-08-15 | 2019-07-30 | Qualcomm Incorporated | Intra video coding using a decoupled tree structure |
JPWO2018047668A1 (en) | 2016-09-12 | 2019-06-24 | ソニー株式会社 | Image processing apparatus and image processing method |
WO2018049594A1 (en) | 2016-09-14 | 2018-03-22 | Mediatek Inc. | Methods of encoder decision for quad-tree plus binary tree structure |
US11095892B2 (en) | 2016-09-20 | 2021-08-17 | Kt Corporation | Method and apparatus for processing video signal |
US10631002B2 (en) * | 2016-09-30 | 2020-04-21 | Qualcomm Incorporated | Frame rate up-conversion coding mode |
US10448010B2 (en) | 2016-10-05 | 2019-10-15 | Qualcomm Incorporated | Motion vector prediction for affine motion models in video coding |
CN109804630A (en) | 2016-10-10 | 2019-05-24 | 夏普株式会社 | The system and method for motion compensation are executed to video data encoding |
US10750190B2 (en) | 2016-10-11 | 2020-08-18 | Lg Electronics Inc. | Video decoding method and device in video coding system |
US20180109810A1 (en) | 2016-10-17 | 2018-04-19 | Mediatek Inc. | Method and Apparatus for Reference Picture Generation and Management in 3D Video Compression |
US11343530B2 (en) | 2016-11-28 | 2022-05-24 | Electronics And Telecommunications Research Institute | Image encoding/decoding method and device, and recording medium having bitstream stored thereon |
CN116886929A (en) | 2016-11-28 | 2023-10-13 | 韩国电子通信研究院 | Method and apparatus for encoding/decoding image and recording medium storing bit stream |
WO2018110203A1 (en) | 2016-12-16 | 2018-06-21 | シャープ株式会社 | Moving image decoding apparatus and moving image encoding apparatus |
US10750203B2 (en) | 2016-12-22 | 2020-08-18 | Mediatek Inc. | Method and apparatus of adaptive bi-prediction for video coding |
US10681370B2 (en) | 2016-12-29 | 2020-06-09 | Qualcomm Incorporated | Motion vector generation for affine motion model for video coding |
US10873744B2 (en) | 2017-01-03 | 2020-12-22 | Lg Electronics Inc. | Method and device for processing video signal by means of affine prediction |
US20190335170A1 (en) | 2017-01-03 | 2019-10-31 | Lg Electronics Inc. | Method and apparatus for processing video signal by means of affine prediction |
US10931969B2 (en) | 2017-01-04 | 2021-02-23 | Qualcomm Incorporated | Motion vector reconstructions for bi-directional optical flow (BIO) |
US20180199057A1 (en) | 2017-01-12 | 2018-07-12 | Mediatek Inc. | Method and Apparatus of Candidate Skipping for Predictor Refinement in Video Coding |
US10701366B2 (en) | 2017-02-21 | 2020-06-30 | Qualcomm Incorporated | Deriving motion vector information at a video decoder |
US10523964B2 (en) | 2017-03-13 | 2019-12-31 | Qualcomm Incorporated | Inter prediction refinement based on bi-directional optical flow (BIO) |
US10701390B2 (en) | 2017-03-14 | 2020-06-30 | Qualcomm Incorporated | Affine motion information derivation |
KR20180107761A (en) | 2017-03-22 | 2018-10-02 | 한국전자통신연구원 | Method and apparatus for prediction using reference block |
US10701391B2 (en) | 2017-03-23 | 2020-06-30 | Qualcomm Incorporated | Motion vector difference (MVD) prediction |
US10873760B2 (en) | 2017-04-07 | 2020-12-22 | Futurewei Technologies, Inc. | Motion vector (MV) constraints and transformation constraints in video coding |
US10805630B2 (en) | 2017-04-28 | 2020-10-13 | Qualcomm Incorporated | Gradient based matching for motion search and derivation |
US20180332298A1 (en) | 2017-05-10 | 2018-11-15 | Futurewei Technologies, Inc. | Bidirectional Prediction In Video Compression |
WO2018212578A1 (en) | 2017-05-17 | 2018-11-22 | 주식회사 케이티 | Method and device for video signal processing |
US10904565B2 (en) | 2017-06-23 | 2021-01-26 | Qualcomm Incorporated | Memory-bandwidth-efficient design for bi-directional optical flow (BIO) |
EP3646598A1 (en) | 2017-06-26 | 2020-05-06 | InterDigital VC Holdings, Inc. | Multiple predictor candidates for motion compensation |
US11172203B2 (en) | 2017-08-08 | 2021-11-09 | Mediatek Inc. | Intra merge prediction |
US10880573B2 (en) | 2017-08-15 | 2020-12-29 | Google Llc | Dynamic motion vector referencing for video coding |
WO2019050115A1 (en) | 2017-09-05 | 2019-03-14 | 엘지전자(주) | Inter prediction mode based image processing method and apparatus therefor |
JP2021005741A (en) | 2017-09-14 | 2021-01-14 | シャープ株式会社 | Image coding device and image decoding device |
US10785494B2 (en) | 2017-10-11 | 2020-09-22 | Qualcomm Incorporated | Low-complexity design for FRUC |
CN109963155B (en) | 2017-12-23 | 2023-06-06 | 华为技术有限公司 | Prediction method and device for motion information of image block and coder-decoder |
WO2020065520A2 (en) | 2018-09-24 | 2020-04-02 | Beijing Bytedance Network Technology Co., Ltd. | Extended merge prediction |
EP3741115A1 (en) | 2018-01-16 | 2020-11-25 | Vid Scale, Inc. | Motion compensated bi-prediction based on local illumination compensation |
US10757417B2 (en) | 2018-01-20 | 2020-08-25 | Qualcomm Incorporated | Affine motion compensation in video coding |
US10687071B2 (en) | 2018-02-05 | 2020-06-16 | Tencent America LLC | Method and apparatus for video coding |
US11012715B2 (en) | 2018-02-08 | 2021-05-18 | Qualcomm Incorporated | Intra block copy for video coding |
US20190306502A1 (en) | 2018-04-02 | 2019-10-03 | Qualcomm Incorporated | System and method for improved adaptive loop filtering |
US10708592B2 (en) | 2018-04-02 | 2020-07-07 | Qualcomm Incorporated | Deblocking filter for video coding and processing |
US20190320181A1 (en) | 2018-04-17 | 2019-10-17 | Qualcomm Incorporated | Generation of motion vector predictors from multiple neighboring blocks in video coding |
US10779002B2 (en) | 2018-04-17 | 2020-09-15 | Qualcomm Incorporated | Limitation of the MVP derivation based on decoder-side motion vector derivation |
US20190364295A1 (en) | 2018-05-25 | 2019-11-28 | Tencent America LLC | Method and apparatus for video coding |
US10986340B2 (en) | 2018-06-01 | 2021-04-20 | Qualcomm Incorporated | Coding adaptive multiple transform information for video coding |
WO2019234598A1 (en) | 2018-06-05 | 2019-12-12 | Beijing Bytedance Network Technology Co., Ltd. | Interaction between ibc and stmvp |
GB2589222B (en) | 2018-06-07 | 2023-01-25 | Beijing Bytedance Network Tech Co Ltd | Sub-block DMVR |
US11303923B2 (en) | 2018-06-15 | 2022-04-12 | Intel Corporation | Affine motion compensation for current picture referencing |
TWI746994B (en) | 2018-06-19 | 2021-11-21 | 大陸商北京字節跳動網絡技術有限公司 | Different precisions for different reference list |
GB2589223B (en) | 2018-06-21 | 2023-01-25 | Beijing Bytedance Network Tech Co Ltd | Component-dependent sub-block dividing |
TWI704803B (en) | 2018-06-29 | 2020-09-11 | 大陸商北京字節跳動網絡技術有限公司 | Restriction of merge candidates derivation |
TWI719519B (en) | 2018-07-02 | 2021-02-21 | 大陸商北京字節跳動網絡技術有限公司 | Block size restrictions for dmvr |
US10362330B1 (en) | 2018-07-30 | 2019-07-23 | Tencent America LLC | Combining history-based motion vector prediction and non-adjacent merge prediction |
WO2020031062A1 (en) | 2018-08-04 | 2020-02-13 | Beijing Bytedance Network Technology Co., Ltd. | Interaction between different dmvd models |
KR20240005178A (en) | 2018-09-19 | 2024-01-11 | 베이징 바이트댄스 네트워크 테크놀로지 컴퍼니, 리미티드 | Syntax reuse for affine mode with adaptive motion vector resolution |
US11212550B2 (en) | 2018-09-21 | 2021-12-28 | Qualcomm Incorporated | History-based motion vector prediction for affine mode |
KR20240000644A (en) | 2018-09-22 | 2024-01-02 | 엘지전자 주식회사 | Method and apparatus for processing video signal based on inter prediction |
CN110944191A (en) | 2018-09-23 | 2020-03-31 | 北京字节跳动网络技术有限公司 | Signaling of motion vector accuracy indication with adaptive motion vector resolution |
US11051034B2 (en) | 2018-10-08 | 2021-06-29 | Qualcomm Incorporated | History-based motion vector predictor |
US11284066B2 (en) | 2018-10-10 | 2022-03-22 | Tencent America LLC | Method and apparatus for intra block copy in intra-inter blending mode and triangle prediction unit mode |
CN112913240A (en) | 2018-10-22 | 2021-06-04 | 北京字节跳动网络技术有限公司 | Collocation between decoder-side motion vector derivation and other codec tools |
WO2020084461A1 (en) | 2018-10-22 | 2020-04-30 | Beijing Bytedance Network Technology Co., Ltd. | Restrictions on decoder side motion vector derivation based on coding information |
WO2020084510A1 (en) | 2018-10-23 | 2020-04-30 | Beijing Bytedance Network Technology Co., Ltd. | Adaptive control point selection for affine coding |
WO2020088691A1 (en) | 2018-11-02 | 2020-05-07 | Beijing Bytedance Network Technology Co., Ltd. | Harmonization between geometry partition prediction mode and other tools |
KR20230158645A (en) | 2018-11-05 | 2023-11-20 | 베이징 바이트댄스 네트워크 테크놀로지 컴퍼니, 리미티드 | Interpolation for inter prediction with refinement |
KR20210089155A (en) | 2018-11-10 | 2021-07-15 | 베이징 바이트댄스 네트워크 테크놀로지 컴퍼니, 리미티드 | Rounding in the fairwise average candidate calculation |
WO2020103934A1 (en) | 2018-11-22 | 2020-05-28 | Beijing Bytedance Network Technology Co., Ltd. | Construction method for inter prediction with geometry partition |
US11032574B2 (en) | 2018-12-31 | 2021-06-08 | Tencent America LLC | Method and apparatus for video coding |
US11115653B2 (en) | 2019-02-22 | 2021-09-07 | Mediatek Inc. | Intra block copy merge list simplification |
WO2020185429A1 (en) | 2019-03-11 | 2020-09-17 | Alibaba Group Holding Limited | Method, device, and system for determining prediction weight for merge mode |
-
2019
- 2019-06-21 CN CN201910544642.0A patent/CN110636298B/en active Active
- 2019-06-21 WO PCT/IB2019/055244 patent/WO2019244117A1/en active Application Filing
- 2019-06-21 TW TW108121836A patent/TWI739120B/en active
-
2020
- 2020-11-16 US US17/099,042 patent/US11197003B2/en active Active
Patent Citations (3)
Publication number | Priority date | Publication date | Assignee | Title |
---|---|---|---|---|
CN106537915A (en) * | 2014-07-18 | 2017-03-22 | 联发科技(新加坡)私人有限公司 | Method of motion vector derivation for video coding |
CN106559669A (en) * | 2015-09-29 | 2017-04-05 | 华为技术有限公司 | The method and device of image prediction |
WO2017118411A1 (en) * | 2016-01-07 | 2017-07-13 | Mediatek Inc. | Method and apparatus for affine inter prediction for video coding system |
Non-Patent Citations (1)
Title |
---|
A unified condition for affine merge and affine inter mode;Jaeho Lee等;《Joint Video Exploration Team (JVET) of ITU-T SG 16 WP 3 and ISO/IEC JTC 1/SC 29/WG 11》;20170104;第1-3页 * |
Also Published As
Publication number | Publication date |
---|---|
US20210352302A1 (en) | 2021-11-11 |
TWI739120B (en) | 2021-09-11 |
TW202007156A (en) | 2020-02-01 |
US20210076050A1 (en) | 2021-03-11 |
US11197003B2 (en) | 2021-12-07 |
CN110636298A (en) | 2019-12-31 |
WO2019244117A1 (en) | 2019-12-26 |
Similar Documents
Publication | Publication Date | Title |
---|---|---|
CN110636298B (en) | Unified constraints for Merge affine mode and non-Merge affine mode | |
US11895306B2 (en) | Component-dependent sub-block dividing | |
US20200396465A1 (en) | Interaction between ibc and affine | |
CN110944204B (en) | Simplified space-time motion vector prediction | |
CN111010571B (en) | Generation and use of combined affine Merge candidates | |
US11805259B2 (en) | Non-affine blocks predicted from affine motion | |
CN110944182A (en) | Motion vector derivation for sub-blocks in affine mode | |
CN110662073B (en) | Boundary filtering of sub-blocks | |
US11968377B2 (en) | Unified constrains for the merge affine mode and the non-merge affine mode |
Legal Events
Date | Code | Title | Description |
---|---|---|---|
PB01 | Publication | ||
PB01 | Publication | ||
SE01 | Entry into force of request for substantive examination | ||
SE01 | Entry into force of request for substantive examination | ||
GR01 | Patent grant | ||
GR01 | Patent grant |