System and method for layered video coding enhancement

Pulse or digital communications – Bandwidth reduction or expansion – Television or motion video signal

Reexamination Certificate

Rate now

  [ 0.00 ] – not rated yet Voters 0   Comments 0

Details

C375S240100, C375S240110

Reexamination Certificate

active

06510177

ABSTRACT:

TECHNICAL FIELD
The present invention relates in general to video compression and more particularly to a system and a method for encoding, transmitting, decoding and storing a high-resolution video sequence using a low-resolution base layer and a higher-resolution enhancement layer.
BACKGROUND OF THE INVENTION
Bit rate reduction is vitally important to achieve the objective of sending as much information as possible with a given communication or storage capacity. Bit rate is the amount of data that can be transmitted in a given time. Bit rate reduction is important because communication capacity is limited by regulatory, physical and commercial constraints and, as demand increases for higher resolution television and video, it is crucial that maximum use be made of the limited capacity available on any given communication or storage medium.
One technique of managing bit rate is data compression. Data compression is storing data in a format that requires less space than would otherwise be used to store the information. Data compression is particularly useful in the transmission of information because it allows a large amount of information to be transmitted using a reduced number of bits. Lossless data compression, which is used mainly for compressing text information, programs, or other computer data refers to data compression formats in which no data is lost. Greater compression can be achieved on graphics, audio and video data by using lossy compression, which refers to data compression formats in which some amount of representation fidelity is lost. Most video compression formats use a lossy compression technique. A compression method with a high degree of bit rate reduction for a given level of fidelity is said to have good compression efficiency.
The well-known International Telecommunications Union-Telecommunications (ITU-T) H.26x and Moving Picture Experts Group (MPEG) video coding standards are examples of a family of conventional video compression formats that use lossy compression. These coding techniques provide high compression rates by representing some image frames as only the changes between frames rather than the entire frame. The changing information is then encoded using a technique called Motion-Compensated Discrete Cosine Transform (MC+DCT) coding. Motion compensation (MC) approximates each area of a video picture as a spatially-shifted area of a previously-decoded picture, and Discrete. Cosine Transform (DCT) coding is a technique that represents waveform data as a weighted sum of cosine waveforms. In general, ITU-T and MPEG video compression remove temporal redundancy between video frames by means of motion compensation, remove spatial redundancy within a video frame by means of a Discrete Cosine Transform and quantization approximation rounding of the DCT samples, and to remove statistical redundancy of quantized index values by means of statistical lossless entropy-reduction coding.
More particularly, ITU-T and MPEG coding work by dividing each frame into rectangular (such as 16×16 pixel) macroblocks and first determining how each macroblock has moved between frames. A motion vector defines any motion of the macroblock that occurs between frames and is used to construct a predicted frame. A process called motion estimation takes place in the encoder to determine the best motion vector value for each macroblock. This predicted frame, which is a previously-decoded frame adjusted by the motion vectors, is compared to an actual input frame. Any new information left over that is new is called the residual and used to construct residual frames.
There are generally three main types of coded pictures in such conventional video coding: (1) intra pictures (I-frames); (2) forward predicted pictures (P-frames); and (3) bi-directional predicted pictures (B-frames). I-frames are encoded as independent pictures with no reference to past or future frames. These frames contain full picture information and can be used to predict other frames. P-frames are encoded relative to the past frames, while B-frames are encoded relative to past frames, future frames or both. ITU-T and MPEG coding use these three types of frames and encoded motion vectors to represent video. This video representation is performed by using I-frames at the start of an independent sequence of pictures and then using P and B frames to encode the remaining pictures in the sequence.
One problem with ITU-T and MPEG coding is that the decoding of high-resolution video requires far greater computational complexity than what is required for lower-resolution video. This means that high-resolution decoders are significantly more expensive than those decoders used for lower resolution video. Delivery of high-resolution video also requires a much higher bit rate than does lower-resolution video. It is therefore highly desirable to provide support for delivery of the same video content as either low-resolution video or as high-resolution video.
One technique of video coding that encodes video using a low-resolution base layer and a higher-resolution enhancement layer is known as spatially-scalable video coding. Spatially-scalable video coding uses a base layer that is decodable as a conventional non-layered video representation at a lower bit rate than an enhancement layer used for the high-resolution video. This allows the base layer to serve lower-capacity receivers while enabling better service for higher-capacity receivers (that receive both the base and enhancement layers). The base layer may also be designed to conform to some prior standard encoding method, in order for the base layer to leverage receivers manufactured to popular and widely-used designs.
One disadvantage, however, of spatially-scalable video coding is that there is a significant loss of compression efficiency for the high-resolution video representation relative to a separate encoding of the high resolution video using the same total bit rate but without the scalability layering structure.
There exists a need, therefore, for a system and a method of encoding, transmitting, decoding and storing a high-resolution video sequence that provides higher compression efficiency than current standard techniques while retaining the advantages of spatially-scalable layered video coding. Such a system and a method would have relevance for HDTV and beyond, and could potentially become a universal video protocol for such widespread use as the Internet, digital video disks (DVD) and new generations of home and commercial video recording devices.
SUMMARY OF THE INVENTION
To overcome the limitations in the prior art as described above and other limitations that will become apparent upon reading and understanding the present specification, the present invention is embodied in a system and a method for transmitting and storing high-resolution video using a low-resolution base layer and a higher-resolution enhancement layer. The present invention uses decoded low-resolution images and additional data from the low-resolution video representation to aid in the decoding of the higher-resolution video. In particular, a preferred embodiment uses motion vector data from the low-resolution video representation to aid in the decoding of the higher-resolution video. The present invention provides high fidelity, uses a minimum amount of bit rate, and can be applied in a manner which allows the low-resolution video to remain backward-compatible with existing standard video compression technology (such as the ITU-T and MPEG standards).
In particular, the present invention is especially well-suited for transmitting and delivering encoded higher-resolution video so that it can be viewed simultaneously in low resolution by a base layer decoder and in enhanced high resolution by an enhancement layer decoder. The present invention divides and encodes a high-resolution video sequence into a lower-resolution base layer and a higher-resolution enhancement layer. The low-resolution base layer, although encoded using a special encoder, can remain, if desired, completely

LandOfFree

Say what you really think

Search LandOfFree.com for the USA inventors and patents. Rate them and share your experience with other people.

Rating

System and method for layered video coding enhancement does not yet have a rating. At this time, there are no reviews or comments for this patent.

If you have personal experience with System and method for layered video coding enhancement, we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and System and method for layered video coding enhancement will most certainly appreciate the feedback.

Rate now

     

Profile ID: LFUS-PAI-O-3050155

  Search
All data on this website is collected from public sources. Our data reflects the most accurate information available at the time of publication.