EP1538567A2

Method and apparatus for scalable video encoding and decoding

Abstract

Disclosed is a scalable video coding algorithm. A method for video coding includes temporally filtering frames in the same sequence to a decoding sequence thereof to remove temporal redundancy, obtaining and quantizing transformation coefficients from frames whose temporal redundancy is removed, and generating bitstreams. A video encoder comprises a temporal transformation unit (10), a spatial transformation unit (20), a quantization unit (30) and a bitstream generation unit (40) to perform the method. A method for video decoding is basically reverse in sequence to the video coding. A video decoder extracts information necessary for video decoding by interpreting the received bitstream and decoding it. Thus, video streams may be generated by allowing a decoder to decode the generated bitstreams, while maintaining the temporal scalability on an encoder-side.

EP1538567A2, drawing sheet 1
Sheet 1 of 18

Term

Term ended

Projected expiry passed 30 November 2024, 1.8 years ago.

  1. Priority
  2. Filed
  3. Published
  4. Projected expiry
  5. Today

23 claims: 11 independent, 12 dependent

  1. 1
    A method for video coding, the method comprising:(a) receiving a plurality of frames constituting a video sequence and sequentially eliminating a temporal redundancy between the plurality of frames on a Group Of Pictures (GOP) basis, starting from a frame at a highest temporal level;and (b) generating a bit-stream by quantizing transformation coefficients obtained from the plurality of frames whose temporal redundancy has been eliminated.
  2. 4
    The method as claimed in any preceding claim, wherein in step (a), the frame at the highest temporal level is set to an A frame when a temporal redundancy between frames constituting a GOP is eliminated, the temporal redundancy between the frames of the GOP other than the A frame at the highest temporal level is eliminated in the sequence from the highest temporal level to a lowest temporal level, and is eliminated in the sequence from a lowest frame index to a highest frame index when the frames are at the same temporal level, where one or more frames which can be referenced by each frame in the course of eliminating the temporal redundancy have a higher index than the frames at a higher temporal level or the same temporal level.
  3. 7
    The method as claimed in any preceding claim, further comprising eliminating spatial redundancy between the plurality of frames, wherein the generated bit-stream further comprises information on a sequence of spatial redundancy elimination and temporal redundancy elimination.
  4. 8
    A video encoder comprising:a temporal transformation unit (10) receiving a plurality of frames and eliminating a temporal redundancy of the frames in a sequence from a highest temporal level to a lowest temporal level;a quantization unit (30) quantizing transformation coefficients obtained after eliminating the temporal redundancy between the frames;and a bit-stream generation unit (40) generating a bit-stream including the quantized transformation coefficients.
  5. 11
    The video encoder as claimed in any one of claims 8 to 10, further comprising a spatial transformation unit (20) eliminating a spatial redundancy between the plurality of frames, wherein the bit-stream generation unit (40) combines information on the sequence for eliminating temporal redundancy and a sequence for eliminating spatial redundancy to obtain the transformation coefficients and generate the bit-stream.
  6. 12
    A method for video decoding, the method comprising:(a) extracting information regarding encoded frames and a redundancy elimination sequence by receiving and interpreting a bit-stream;(b) obtaining transformation coefficients by inverse-quantizing the information regarding the encoded frames;and (c) restoring the encoded frames through an inverse-spatial transformation and an inverse-temporal transformation of the transformation coefficients to the redundancy elimination sequence.
  7. 14
    A video decoder comprising:a bit-stream interpretation unit (100) interpreting a received bit-stream to extract information regarding encoded frames therefrom and a redundancy elimination sequence;an inverse-quantization unit (210) inverse-quantizing the information regarding the encoded frames to obtain transformation coefficients therefrom;an inverse spatial transformation unit (220) performing an inverse-spatial transformation process;and an inverse temporal transformation unit (230) performing an inverse-temporal transformation process,    wherein the encoded frames of the bit-stream are restored by performing the inverse-spatial transformation process and the inverse-temporal transformation process on the transformation coefficients to the redundancy elimination sequence of the encoded frames by referencing the redundancy elimination sequence.
  8. 15
    A storage medium having recorded thereon a program readable by a computer to execute a video coding method, said method comprising:(a) receiving a plurality of frames constituting a video sequence and sequentially eliminating a temporal redundancy between the plurality of frames on a Group Of Pictures (GOP) basis, starting from a frame at a highest temporal level;and (b) generating a bit-stream by quantizing transformation coefficients obtained from the plurality of frames whose temporal redundancy has been eliminated.
  9. 18
    The storage medium of any one of claims 15 to 17, wherein in step (a), the frame at the highest temporal level is set to an A frame when a temporal redundancy between frames constituting a GOP is eliminated, the temporal redundancy between the frames of the GOP other than the A frame at the highest temporal level is eliminated in the sequence from the highest temporal level to a lowest temporal level, and is eliminated in the sequence from a lowest frame index to a highest frame index when the frames are at the same temporal level, where one or more frames which can be referenced by each frame in the course of eliminating the temporal redundancy have a higher index than the frames at a higher temporal level or the same temporal level.
  10. 21
    The storage medium of any one of claims 18 to 20, said method further comprising eliminating spatial redundancy between the plurality of frames, wherein the generated bit-stream further comprises information on a sequence of spatial redundancy elimination and temporal redundancy elimination.
  11. 22
    A storage medium having recorded thereon a program readable by a computer to execute a video decoding method, said method comprising:(a) extracting information regarding encoded frames and a redundancy elimination sequence by receiving and interpreting a bit-stream;(b) obtaining transformation coefficients by inverse-quantizing the information regarding the encoded frames;and (c) restoring the encoded frames through an inverse-spatial transformation and an inverse-temporal transformation of the transformation coefficients to the redundancy elimination sequence.