Video coding

Patent No. US6968005 (titled "Video coding") on May 15, 2001. The application was issued on Nov 22, 2005.

What is this patent about?

’005 is related to the field of video compression and error resilience, specifically addressing the challenges of transmitting temporally predicted video data over unreliable networks. In standard video coding, such as H.263 or MPEG, frames are often predicted from previous reference pictures to reduce redundancy. However, if a reference picture is lost during transmission, the decoder may continue to decode subsequent frames using the wrong anchor, leading to significant temporal error propagation that persists until the next independent INTRA frame is received.

The underlying idea behind ’005 is to decouple the tracking of reference frames from the standard display-order timeline by implementing a dedicated numbering scheme for anchor pictures. By assigning a Reference Picture Order Number (RPON) only to those frames that serve as anchors for future prediction, the system creates a continuous sequence for the decoder to monitor. This allows the decoder to immediately distinguish between the loss of a non-essential B-frame and the loss of a critical reference frame, regardless of how many frames were skipped or reordered during the encoding process.

The claims of ’005 focus on an encoding and decoding method that utilizes a sequence indicator with an independent numbering scheme specifically for reference pictures. Unlike the Temporal Reference (TR) which tracks display timing, this indicator increments by a predetermined amount for every INTRA or INTER frame used as a reference, remaining unaffected by the presence of non-reference pictures. The independent claims cover the generation of this indicator at the encoder, its inclusion in the bitstream, and the decoder’s logic for comparing indicators of consecutive reference frames to detect gaps.

In practice, the invention works by inserting the RPON into the video bitstream, often within the Supplemental Enhancement Information (SEI) or picture headers. When the decoder receives a new reference frame, it subtracts the previous reference frame's RPON from the current one. If the difference exceeds the expected increment, the decoder knows a reference picture was lost. This triggers a specific error-handling protocol, such as sending an immediate request to the transmitter for a new INTRA-coded frame to reset the prediction chain and clear visual artifacts.

This approach differs from prior solutions that relied on transport-level sequence numbers or display-order timestamps, which are often unreliable for detecting reference-specific losses in complex coding structures. By providing a robust continuity check within the video syntax itself, the invention enables decoders to maintain synchronization even when using multi-layer scalability or variable encapsulation strategies. This ensures that real-time applications like video conferencing can recover from network jitter and packet loss without unnecessarily freezing the display for non-critical data gaps.

How does this patent fit in bigger picture?

Technical Landscape

In the early 2000s when ’005 was filed, video compression was typically implemented using hierarchical bit-stream structures where temporal redundancy was reduced through the use of anchor pictures, such as I-frames and P-frames. At a time when systems commonly relied on implicit reference picture signaling, the bit-stream syntax generally lacked explicit identifiers for the specific frames used for motion compensation. When hardware and software constraints made reliable transmission over lossy networks non-trivial, decoders often struggled to maintain synchronization during packet loss. Because transport-level sequence numbers were decoupled from the underlying video syntax, decoders had no native mechanism to distinguish between the loss of a critical reference frame and a non-reference frame, such as a B-picture, often resulting in unnecessary display freezes or the requirement for a full intra-frame refresh.

Prosecution Position

The disclosed invention represents a technical advancement by introducing an explicit temporal order indicator specifically for reference pictures within the video bit-stream. This architectural shift moves away from implicit frame tracking by associating a sequence-based indicator, such as a reference picture order number, with each frame that serves as a temporal prediction anchor. This integration enables a decoder to independently detect the loss of a reference frame by identifying discontinuities in the indicator sequence, even when transport-layer information is missing or inconsistent. The technical effect achieved is the ability to differentiate between critical reference data loss and non-critical picture loss, allowing the system to maintain decoding continuity for non-reference frames and overcome the constraint of total display freezing during minor transmission errors.

Claims

The patent contains a total of 46 claims, with claims 1, 5, 7, 9, and 37 serving as the independent claims. These independent claims focus on a video coding system that utilizes a sequence indicator with an independent numbering scheme to track the encoding order of reference pictures, specifically ensuring that consecutive reference pictures are assigned values that differ by a predetermined amount regardless of any intervening non-reference pictures to facilitate error detection. The dependent claims generally serve to specify the predetermined value of the indicator, define its placement within various headers or bitstream layers, adapt the system for scalable video coding and specific standards like H.263, and describe the integration of the technology into portable radio communications and multimedia terminal devices.

Key Claim Terms New

Definitions of key terms used in the patent claims.

Term (Source)Support for SpecificationInterpretation
Encoding order
(Claim 1, Claim 5, Claim 7, Claim 37)
For example, a section of video bit-stream may contain P-picture P1, B-picture B2, P-picture P3, and P-picture P4, captured (and to be displayed) in this order. However, this section would be compressed, transmitted, and decoded in the following order: P1, P3, B2, P4 since B2 requires both P1 and P3 before it can be encoded or decoded.The specific sequence in which pictures are compressed and arranged in the bit-stream, which may differ from the capture or display order due to temporal prediction requirements.
Independent numbering scheme
(Claim 1, Claim 5, Claim 7, Claim 9, Claim 37)
The invention also enables a decoder to differentiate B picture losses from reference picture losses. Consequently, decoders can continue decoding after a B picture loss instead of waiting for the next INTRA picture. The indicator is incremented by one from the previous reference picture.A numbering method for reference pictures where the assigned values change by a fixed amount regardless of how many non-reference pictures (such as B-pictures) are located between them.
Non-reference pictures
(Claim 1, Claim 5, Claim 7, Claim 9, Claim 37)
B-pictures are not used as anchor pictures, i.e., other pictures are not predicted from them. Therefore they can be discarded (intentionally or unintentionally) without impacting the picture quality of future pictures.Pictures, typically B-pictures, that are predicted from other frames but are not themselves used as a reference for predicting subsequent pictures in the sequence.
Reference pictures
(Claim 1, Claim 5, Claim 7, Claim 9, Claim 37)
Compressed pictures that do not utilise temporal redundancy reduction methods are usually called INTRA or I-frames or I-pictures. Temporally predicted images are usually forwardly predicted from a picture occurring before the current picture and are called INTER or P-frames. B-pictures are inserted between anchor picture pairs of I- and/or P-frames and are predicted from either one or both of these anchor pictures.Pictures within a video sequence, specifically INTRA pictures or certain temporally predicted pictures, that serve as the basis for the temporal prediction of other pictures.
Sequence indicator
(Claim 1, Claim 5, Claim 7, Claim 9, Claim 37)
Thus each reference picture (e.g. I-frames and P-frames) is associated with a sequence number. Preferably the indicator is incremented each time a reference picture is encoded. Most advantageously the indicator is incremented by one each time a reference picture is encoded.A value or sequence number associated with each reference picture to indicate its temporal order within the encoded video signal relative to other reference pictures.
Temporally independent INTRA pictures
(Claim 1, Claim 5, Claim 7, Claim 9, Claim 37)
Compressed pictures that do not utilise temporal redundancy reduction methods are usually called INTRA or I-frames or I-pictures. Since the compression efficiency in INTRA pictures is normally lower than in INTER pictures, INTRA pictures are used sparingly, especially in low bit-rate applications. INTRA pictures are typically inserted to stop temporal propagation of transmission errors in a reconstructed video signal and to provide random access points to a video bit-stream.Compressed video frames that do not rely on information from other pictures for their reconstruction, used to stop error propagation and provide random access points.

Litigation Cases New

US Latest litigation cases involving this patent.

Case NumberFiling DateTitle
0:24-cv-04269Nov 25, 2024Element Television Company, Llc V. Nokia Corporation
1:23-cv-01236Oct 31, 2023Nokia Technologies Oy V. Amazon.Com, Inc.

Patent Family

Patent Family

File Wrapper

The dossier documents provide a comprehensive record of the patent's prosecution history - including filings, correspondence, and decisions made by patent offices - and are crucial for understanding the patent's legal journey and any challenges it may have faced during examination.

  • Get instant alerts for new documents

US6968005

SEP
Application Number
US09855640A
Filing Date
May 15, 2001
Status
Expired
Publication Date
Nov 22, 2005
External Links
Slate, USPTO , Google Patents