Significance map encoding and decoding using partition selection

Patent No. US10448052 (titled "Significance map encoding and decoding using partition selection") on Sep 22, 2017. The application was issued on Oct 15, 2019.

What is this patent about?

’052 is related to the field of video data compression, specifically the entropy encoding and decoding of significance maps. In modern codecs like HEVC, significance maps identify the locations of non-zero coefficients within a transform unit after spectral transformation and quantization. Because these maps account for a large portion of the total bitstream, efficient context-adaptive binary arithmetic coding is essential for reducing data redundancy while maintaining high processing speeds.

The underlying idea behind ’052 is that traditional uniform partitioning of significance maps—where contexts are distributed evenly across a block—fails to account for the statistical concentration of energy in the low-frequency regions. The invention recognizes that bit positions in the upper-left area of a transform unit are used more frequently and carry different probability characteristics than those in the bottom-right. By employing non-spatially-uniform partitioning, the system can allocate more granular contexts to high-traffic areas and shared contexts to low-traffic areas, optimizing the balance between model accuracy and the speed of probability estimation convergence.

The claims of ’052 focus on a specific block-based mapping for 4×4 transform units that assigns nine distinct contexts across the fifteen relevant bit positions. This mapping is defined by a precise numerical sequence—0, 1, 2, 3, 4, 5, 2, 3, 6, 6, 7, 7, 8, 8, 7—which dictates how each coordinate in the block is associated with a context for entropy coding. This specific arrangement ensures that the most critical low-frequency coefficients receive dedicated or carefully grouped contexts, while higher-frequency positions share contexts to prevent the overhead of tracking underutilized probability models.

In practice, the invention operates by looking up the assigned context for each bit position during the scanning process. As the encoder or decoder traverses the block, it retrieves the current probability state for the assigned context, processes the bit, and immediately updates that state to reflect the new data. This context-adaptive approach allows the codec to learn the local statistics of the video slice dynamically. The implementation also supports switching between different partition sets based on the size of the encoded slice, ensuring that the complexity of the context model scales with the amount of data available to train it.

This approach differs from prior solutions by moving away from rigid, uniform grids that treat all coefficient positions with equal weight. By utilizing a refinement-based hierarchy, the invention allows a decoder to initialize a complex partition set using the states of a simpler, coarser set, facilitating a smooth transition as more data is processed. This mechanism effectively solves the problem of context dilution, where having too many contexts for too little data leads to poor probability estimation, thereby improving overall compression efficiency without significantly increasing the computational footprint of the entropy engine.

How does this patent fit in bigger picture?

Technical Landscape

In the early 2010s when ’052 was filed, video compression systems were transitioning toward higher resolution formats at a time when significance map encoding was typically implemented using fixed, position-dependent context models for transform units. When systems commonly relied on tracking a large number of distinct contexts for different block sizes, the memory overhead and computational complexity of managing these context states became a significant bottleneck. During this era, hardware and software constraints made the high-speed lookup and updating of nearly a hundred different probability models non-trivial, particularly as transform unit sizes increased to accommodate high-definition content.

Prosecution Position

The disclosed invention achieves a technical advancement by reducing the memory and processing requirements of entropy coding through a partitioned context selection architecture. Instead of maintaining unique contexts for every coefficient position or small fixed sub-block, the system implements a method of partitioning a transform unit into a plurality of regions and assigning a single, shared context to all coefficient positions within a specific region. This architectural shift enables the encoder and decoder to significantly reduce the total number of contexts tracked—such as reducing the requirements for large transform units to a small set of regional contexts—without sacrificing the statistical accuracy needed for effective compression. The resulting technical effect is a streamlined significance map processing stage that maintains high throughput and lower hardware complexity while handling large-scale residual data.

Claims

This patent contains 21 claims, with independent claims 1, 7, and 13 focusing on a method, an encoder, and a processor-readable medium for encoding a significance map in a 4x4 transform unit by assigning specific context values to bit positions based on a predefined block-based mapping. The dependent claims serve to further refine these processes by specifying the selection of partition sets based on luma or chroma text types, the use of selection information within slice or sequence headers, and the implementation of threshold-based switching to refined partition sets during the encoding of a slice.

Key Claim Terms New

Definitions of key terms used in the patent claims.

Term (Source)Support for SpecificationInterpretation
Block-based mapping
(Claim 1, Claim 7, Claim 13)
The partition set assigns contexts to bit positions in accordance with a block-based mapping given by: 0, 1, 2, 3, 4, 5, 2, 3, 6, 6, 7, 7, 8, 8, 7. These integers represent the contexts assigned to the bit positions of a 4x4 block significance map.A specific spatial arrangement or sequence of integers that defines which context index is applied to each coordinate within a 4x4 block.
Context
(Claim 1, Claim 7, Claim 13)
The entropy encoding of the symbols in significance map is based upon a context model. The encoder and decoder track a total of 30 separate contexts for 4x4 luma and chroma TUs. This means the encoder and decoder keep track of and look up 62 different contexts during the encoding and decoding of the significance map.A probability model or state used in entropy encoding that is associated with a specific position or group of positions in a significance map and is updated based on encoded bit values.
Partition set
(Claim 1, Claim 7, Claim 13)
The 8x8 TUs are partitioned (conceptually for the purpose of context association) into 2x2 blocks such that one distinct context is associated with each 2x2 block in the 8x8 TU. The encoder and decoder track a total of 30 (excluding the bottom right corner positions) separate contexts for 4x4 luma and chroma TUs. The partition set assigns contexts to bit positions in accordance with a block-based mapping.A grouping or mapping scheme that assigns specific contexts to different bit positions within a transform unit to facilitate entropy encoding.
Significance map
(Claim 1, Claim 7, Claim 13)
The significance map indicates the positions in the block (other than the last significant coefficient position) that contain non-zero coefficients. The entropy encoding of the symbols in significance map is based upon a context model. In the case of a 4x4 luma or chroma block or transform unit (TU), a separate context is associated with each coefficient position in the TU.A map of bit values corresponding to coefficient positions in a transform unit, where each bit indicates whether the coefficient at that position is non-zero.
Transform unit
(Claim 1, Claim 7, Claim 13)
The block or matrix of quantized transform domain coefficients is sometimes referred to as a 'transform unit'. When spectrally transforming residual data, many of these standards prescribe the use of a discrete cosine transform (DCT) or some variant thereon. The resulting DCT coefficients are then quantized using a quantizer to produce quantized transform domain coefficients, or indices.A block or matrix of quantized transform domain coefficients, such as those resulting from a discrete cosine transform (DCT) applied to residual data.

Litigation Cases New

US Latest litigation cases involving this patent.

Case NumberFiling DateTitle
1:25-cv-00967Jun 23, 2025Velos Media, LLC v. ByteDance Ltd et al

Patent Family

Patent Family

File Wrapper

The dossier documents provide a comprehensive record of the patent's prosecution history - including filings, correspondence, and decisions made by patent offices - and are crucial for understanding the patent's legal journey and any challenges it may have faced during examination.

  • Get instant alerts for new documents

US10448052

SEP
Application Number
US15712640A
Filing Date
Sep 22, 2017
Status
Granted
Publication Date
Oct 15, 2019
External Links
Slate, USPTO , Google Patents