Video coding

Patent No. US8144764 (titled "Video coding") on Oct 5, 2005. The application was issued on Mar 27, 2012.

What is this patent about?

’764 is related to the field of video compression and error resilience, specifically addressing the challenges of temporal predictive coding in unreliable network environments. In standard video coding, such as H.263 or MPEG, pictures are often predicted from previous frames to reduce redundancy. However, if a reference frame is lost during transmission, the decoder may continue to decode subsequent frames using the wrong reference data, leading to persistent visual artifacts and temporal error propagation that lasts until the next independent INTRA frame is received.

The underlying idea behind ’764 is to decouple the tracking of reference frames from the standard temporal display order by implementing an independent numbering scheme. While traditional temporal references track the timing of every frame (including those not used for prediction), this invention introduces a dedicated sequence indicator that only increments for frames that serve as anchors for future prediction. By focusing exclusively on the sequence of reference pictures, the system provides a robust mechanism for a decoder to verify the integrity of its prediction chain regardless of how many non-reference frames were skipped or lost.

The claims of ’764 focus on a method and apparatus for assigning consecutive sequence indicator values to reference pictures in their encoding order. This numbering is strictly independent of the number of non-reference pictures (like B-frames) or non-coded pictures situated between them. The independent claims specifically cover the logic where a decoder identifies a gap in these values—detecting if the difference exceeds a predetermined amount—to immediately flag the loss or corruption of a critical reference frame, rather than relying on external transport-layer sequence numbers.

In practice, the encoder assigns a Reference Picture Order Number (RPON) to each I-frame or P-frame and embeds this value within the video bitstream, often utilizing Supplemental Enhancement Information (SEI) or picture headers. When the decoder receives a new reference frame, it compares the current RPON to the previously stored value. If the increment is greater than the expected step (typically one), the decoder instantly recognizes that a reference dependency has been broken. This allows the decoder to take proactive measures, such as freezing the display or signaling the encoder to transmit a new INTRA-coded recovery point.

This approach differentiates itself from prior art by solving the ambiguity inherent in standard temporal references and transport-layer sequence numbers. Traditional temporal references change based on the display timing, making it difficult to distinguish between a skipped non-reference frame and a lost reference frame. By using a dedicated reference counter, ’764 ensures that the decoder can maintain synchronization of the reference buffer even when variable encapsulation strategies or multi-layer scalability are employed, significantly improving visual stability in error-prone streaming and conferencing applications.

How does this patent fit in bigger picture?

Technical Landscape

In the early 2000s when ’764 was filed, video compression systems were typically implemented using hierarchical bit-stream structures that differentiated between independent intra-coded frames and temporally predicted inter-coded frames. At a time when systems commonly relied on the implicit assumption that the decoder would maintain synchronization with the encoder's reference frame buffer, the standard bit-stream syntax often omitted explicit identification of reference pictures. When hardware or software constraints made high-bandwidth transmission non-trivial, especially in low bit-rate or wireless environments, the reliance on variable length coding and temporal prediction meant that data loss frequently led to sustained error propagation. During this era, while transport-level protocols might provide packet sequence numbers, these were not intrinsically mapped to the video decoding order, making it difficult for a decoder to distinguish between the loss of a non-essential bi-directionally predicted frame and a critical reference frame.

Prosecution Position

The disclosed invention addresses the technical problem of undetected reference picture loss and the resulting temporal error propagation in video decoding. The architectural solution involves the integration of a specific indicator, such as a reference picture order number, directly into the encoded video signal for every picture that serves as a temporal prediction reference. This represents a shift from implicit buffer management to an explicit signaling mechanism where the indicator is incremented according to the temporal order of reference pictures in the bit-stream. The technical effect achieved is the enabling of the decoder to independently verify the continuity of the reference chain regardless of the transport protocol used. This capability allows a decoder to distinguish between the loss of disposable frames and essential reference frames, thereby overcoming the constraint of unnecessary display freezes and allowing for more precise error recovery requests or concealment actions.

Claims

The patent contains a total of 62 claims, with independent claims 1, 15, 31, and 46 directed to methods and apparatuses for encoding and decoding video signals using an independent numbering scheme. These independent claims focus on assigning consecutive reference pictures sequence indicator values that differ by a predetermined amount regardless of the presence of non-reference or non-coded pictures, allowing a decoder to detect corruption or loss by identifying deviations from this expected difference. The dependent claims further specify implementation details such as incrementing values by one, placing indicators in picture or macroblock headers, applying the scheme to multi-layer coding or specific formats like H.263, and incorporating the technology into portable radio communications or multimedia terminal devices.

Key Claim Terms New

Definitions of key terms used in the patent claims.

Term (Source)Support for SpecificationInterpretation
Consecutive reference pictures
(Claim 1, Claim 15, Claim 31, Claim 46)
Compressed pictures that do not utilise temporal redundancy reduction methods are usually called INTRA or I-frames or I-pictures. Temporally predicted images are usually forwardly predicted from a picture occurring before the current picture and are called INTER or P-frames. B-pictures are not used as anchor pictures, i.e., other pictures are not predicted from them.Successive pictures in the encoding or decoding order that serve as temporal anchors (such as I-frames or P-frames) for other pictures, excluding non-anchor frames like B-pictures.
Independent numbering scheme
(Claim 1, Claim 15, Claim 31, Claim 46)
The invention enables a decoder to differentiate B picture losses from reference picture losses. Including this indicator means that a decoder is capable of determining whether a reference picture has been lost and to take appropriate action, if available. This is the case even if the transport protocol does not include sequence information about the packets being transmitted or the transmitter uses a varying encapsulation strategy.A sequence tracking mechanism for reference pictures that operates separately from the total count of all pictures (such as B-pictures or non-coded frames) in a video stream.
Non-reference pictures
(Claim 1, Claim 15, Claim 31, Claim 46)
B-pictures are not used as anchor pictures, i.e., other pictures are not predicted from them. Therefore they can be discarded (intentionally or unintentionally) without impacting the picture quality of future pictures. The invention also enables a decoder to differentiate B picture losses from reference picture losses.Frames within a video sequence, such as B-pictures, that are not used as anchors for the temporal prediction of other frames and can be discarded without affecting future picture quality.
Predetermined amount
(Claim 1, Claim 15, Claim 31, Claim 46)
Preferably the indicator is incremented by one from the previous reference picture. If multi-layer coding is used, preferably this indicator is incremented by one from the previous reference picture in the same enhancement layer.A fixed numerical increment (typically one) used to advance the sequence indicator between one reference picture and the next in encoding order.
Sequence indicator values
(Claim 1, Claim 15, Claim 31, Claim 46)
Each reference picture (e.g. I-frames and P-frames) is associated with a sequence number. Preferably the indicator is incremented each time a reference picture is encoded. Most advantageously the indicator is incremented by one each time a reference picture is encoded.Numerical markers or sequence numbers assigned to reference pictures to denote their specific temporal order within the encoded bit-stream.

Litigation Cases New

US Latest litigation cases involving this patent.

Case NumberFiling DateTitle
0:24-cv-04269Nov 25, 2024Element Television Company, Llc V. Nokia Corporation
1:23-cv-01236Oct 31, 2023Nokia Technologies Oy V. Amazon.Com, Inc.

Patent Family

Patent Family

File Wrapper

The dossier documents provide a comprehensive record of the patent's prosecution history - including filings, correspondence, and decisions made by patent offices - and are crucial for understanding the patent's legal journey and any challenges it may have faced during examination.

  • Get instant alerts for new documents

US8144764

SEP
Application Number
US11242888A
Filing Date
Oct 5, 2005
Status
Expired
Publication Date
Mar 27, 2012
External Links
Slate, USPTO , Google Patents