A coding system for video signals.
Abstract
A video processing system is disclosed that separates and separately encodes and decodes the low (LFREQ) and high (HFREQ) spatial frequency coefficients of images for transmission or storage. Each block of an image is transformed (in 106) into the frequency domain. High frequency coefficients of the resulting transform matrix are separated (in 108) from the low frequency coefficients. The low frequency coefficients are motion prediction compensated (by 124) to derive motion vectors and a prediction error signal. The motion vectors, prediction error signal and high frequency coefficients are channel encoded (in 116) for storage or transmission. In a receiver, the motion vectors and prediction error signal are used to reconstruct a low frequency motion-compensated version of the image. The high frequency coefficients are inverse transformed into the pel domain and are combined with the reconstructed low frequency version of the image to reconstruct a version of the original image.

Term
Term ended
Projected expiry passed 18 December 2010, 15.8 years ago.
- Priority
- Filed
- Published
- Projected expiry
- Today
10 claims: 2 independent, 8 dependent
- 1Apparatus for use in a video coding system, CHARACTERISED BY means for receiving as an input an original digitized video image signal, means (104) for delaying said video image signal, means (106) for deriving frequency transform coefficients from said delayed video image signal, means (108) for separating said frequency transform coefficients into a set of low frequency coefficients and a set of high frequency coefficients, said separation being based upon a predetermined threshold, means (112,118,120,122,124,126) responsive to said set of low frequency coefficients and said input video image signal for generating and supplying as an output a representation of a motion compensated low frequency version of said original video image signal that includes a set of motion vectors and a prediction error signal, means (116) for encoding said set of high frequency coefficients, said prediction error signal and said set of motion vectors into a predetermined format specific to a predetermined output medium, and means for interfacing said predetermined format to said predetermined output medium.
- 8Apparatus for use in a video decoding system, CHARACTERISED BY means (202) for interfacing to a predetermined input medium for receiving, separating, decoding and supplying as an output a set of high frequency coefficients, a prediction error signal and a set of motion vectors from data supplied from said medium in a predetermined format, means (204) for transforming said set of high frequency coefficients into a high frequency pel domain version of an original video image, means (206,208,212,214) responsive to said prediction error signal and said set of motion vectors for deriving a motion compensated low frequency pel domain version of said original video image signal, means (210) for combining said high frequency pel domain version of said original video image and said motion compensated low frequency pel domain version of said original video image signal into a reconstructed full frequency version of said original video image signal in the pel domain, and means for supplying said reconstructed full frequency version of said original video image signal in the pel domain as an output.
Independent claims2
31 paragraphs, as filed
<u style="single">Technical</u>
<b>Field</b>
0001This invention relates to signal coding systems and, more particularly, to encoding and decoding video signals of moving images suitable for transmission or for storage.
<u style="single">Background</u>
<b>of the Invention</b>
0002Prior video signal interframe coding, transmission and reproduction systems initially encode a base frame image. This base frame image incorporates both the low and the high spatial frequency components of the image. The base frame image is transmitted and thereafter only information representative of the difference between the base frame and each subsequent image is transmitted. Both low and high frequency spatial components are contained within the difference information. Using the base frame image and the subsequent difference information a high quality version, i.e., an accurate or error free representation, of any desired subsequent image can be reconstructed. If difference information is lost during transmission, only a low quality version of the desired image, i.e., an inaccurate or error strewn representation, can be reconstructed. Such prior arrangements are inefficient because they lack the ability to rapidly reproduce high quality images in random access applications. Similarly, this prior arrangement requires a long time to recover from loss of the difference information caused by transmission errors. The inefficiency and delays are caused by the leveraged encoding of each image beyond the base frame since all of the information prior to the selected frame is required, i.e., all the information contained in the base frame and all the difference information relating to each sequentially subsequent image. However, it is known that interframe coding tends to reduce the number of bits that are required to be transmitted.
0003In a prior attempt to reduce the reconstruction difficulty a new base frame is periodically incorporated into the bit stream However, this technique dramatically increases the average number of bits required to represent all the information in an average frame because each base frame is encoded using interframe coding and all of the information representing a base frame is used by each subsequent image. It is known that intraframe coding tends to require a larger number of bits than interframe coding. In intraframe coding, characteristics of neighboring picture elements (pels or pixels) are predicted based on the values of those characteristics of neighboring pels in the same frame, and the error or difference between the actual value and the predicted value is encoded. This type of encoding is illustrated in an article entitled "Adaptive Coding of Monochrome and Color Images", IEEE Trans. Communications, Vol. COM-25, pp. 1285-1292, Nov. 1977.
<u style="single">Summary</u>
<b>of the Invention</b>
0004The problems with prior image coding and reproduction systems are overcome, in accordance with an aspect of the invention, by encoding and transmitting a low spatial frequency representation of an image using interframe prediction techniques. High spatial frequency coefficients of the image are encoded directly for transmission.
0005In a specific embodiment of an encoder, a video signal is divided into blocks comprising an array of pels. Each block is transformed into the frequency domain. The high and the low frequency coefficients of each block are separately extracted from the resulting transformed matrix. A version of the original image represented by the low frequency coefficients of each block is motion compensated to yield a prediction error signal and motion vectors. The prediction error signal, the motion vectors and the high frequency coefficients are encoded using a method suitable to the medium of transmission or storage. The prediction error signal and motion vectors are assigned a high priority. The high frequency coefficients are assigned a low priority and may be dropped by a network during congestion or not retrieved from storage.
0006A receiver decodes and separates the prediction error signal, the motion vectors and the high frequency coefficients. The prediction error signal and motion vectors are used to reconstruct a motion compensated version of the low frequency image. The high frequency coefficients are inverse transformed and the resulting reconstructed high frequency version of the image is combined with the reconstructed, motion compensated low frequency version of the image to produce a reconstructed version of the original image.
<u style="single">Brief</u>
<b>Description of the Drawing</b>
0007<ul id="ul0001" list-style="none"><li>FIG. 1 shows, in simplified block diagram form, an encoder embodying aspects of the invention; and</li><li>FIG. 2 shows, in simplified block diagram form, a decoder embodying aspects of the invention.</li></ul>
<u style="single">Detailed</u>
<b>Description</b>
0008There is some experimental evidence that for regions of high motion in a sequence of moving image frames the interframe correlation of high frequency coefficients of a discrete cosine transform (DCT) of those regions is approximately zero (0). Therefore, the bit-rate to code these coefficients in an intraframe format is not much different than coding them in an interframe frame or Motion Compensated format. FIG. 1 shows an encoder that encodes the high frequency coefficients of a DCT of a sequence of image frames in an intraframe format. An additional DCT beyond those used in prior motion compensated systems is used as a low-pass loop-filter to generate a low frequency prediction.
0009In an example encoder, an original video signal VIDIN is first delayed by delay 102 for reasons to be given later. The delayed signal 104 is then converted from the pel domain to the Discrete Cosine Transform domain by DCT 106. An example DCT 106 groups input pels, i.e. picture elements, into 2- dimensional blocks, e.g., 8x8 pels each. DCT 106 produces a set of output frequency coefficients that are grouped into 2-dimensional blocks of the same size as the input blocks. The transform coefficients supplied as an output by DCT 106 are then partitioned by separator 108 into two groups, namely, high spatial frequency coefficients and low spatial frequency coefficients. The high spatial frequency coefficients, HFREQ, are selected based upon a predetermined threshold and are supplied to quantizer 110. The remaining low spatial frequency coefficients, LFREQ, are supplied to subtracter 112.
0010High spatial frequency coefficients HFREQ are quantized by quantizer 110. A quantizer reduces the number of levels available for coefficients to assume. The quantizing can optionally be made responsive to the quantity of information stored in elastic store 114 so as to prevent elastic store 114 from overflowing. The output from quantizer 110, QHFREQ, is supplied to elastic store 114. Elastic store 114 subsequently supplies QHFREQ to channel coder 116 for encoding appropriate for transmission or storage, depending on the application. If elastic store 114 is not present, QHFREQ is supplied directly to channel coder 116 for encoding. Low spatial frequency coefficients LFREQ are coded by interframe prediction techniques that are well known in the art and exploit the correlation normally present between image frames. For example, here we show simple Motion Compensation prediction. (See Digital Pictures Representation and Compression by Arun N. Netravali and Barry G. Haskell, Plenum Press 1988 pp. 334-340.) Other methods might include Conditional Motion Compensated Interpolation (see co-pending application Serial Number 413,520 Filed on September 27, 1989 and allowed on April 5, 1990).
0011Low frequency coefficient prediction signal PRED, supplied to subtracter 112, is subtracted from the low frequency coefficients LFREQ to yield an initial low frequency prediction error signal LPERR. Low frequency prediction error signal LPERR is quantized by quantizer 118 and then supplied as low frequency prediction error signal QLPERR to elastic store 114. In turn, QLPERR is subsequently encoded and supplied for transmission by channel coder 116. Data supplied to channel coder 116 from elastic store 114 is encoded using well known reversible data compression methods. A serial bit-stream output that can be transmitted or stored is supplied as an output from channel coder 116.
0012It should be noted that in some applications quantizers 110 and 118 and elastic store 114 may be eliminated. If quantizer 118 is eliminated, low frequency prediction error signal LPERR is supplied directly to elastic store 114 and adder 120. Additionally, a significance detector (not shown) may be incorporated into quantizer 118. This detector would determine whether the values of the low frequency coefficients in a block of QLPERR were too small to warrant transmission. Insignificant blocks would be replaced by zero (0) before being supplied as an output.
0013After any significance thresholding is performed, the quantized prediction error signal QLPERR is added to low frequency coefficient prediction signal PRED by adder 120 to produce "reconstructed" low frequency coefficients RLFREQ. These reconstructed coefficients are inverse transformed by INVERSE DCT 122 to form a low frequency pel domain signal LPEL. LPEL is passed into a frame memory 123 of motion compensation prediction unit 124. Motion compensation prediction units are well known in the art. Original signal VIDIN is also supplied to motion compensation prediction unit 124. Delay 102 is required to insure that coefficients LFREQ correspond in time to predicted coefficients PRED due to the processing delay in motion compensation prediction unit 124. Motion compensation prediction unit 124 compares each block of the original signal VIDIN with signal LPEL stored in its frame memory 123 and calculates a motion vector for each block. Motion vectors are necessary to reconstruct a shift-matrix for the motion compensated low frequency version of the image. The motion vectors are supplied as output signal MOVECT to elastic store 114 for subsequent transmission by channel coder 116. Channel coder 116 assigns a higher priority to the low frequency quantized prediction error signal and the motion vectors than is assigned to the high frequency coefficients.
0014Motion compensation prediction unit 124 also generates MCPEL, which is a block of pels that is shifted by an amount corresponding to each of the aforementioned motion vectors. MCPEL represents the prediction to be used in coding the corresponding block of original pels now being supplied as an output by delay 102. Thus, an original block of pels on signal line 104 passes into DCT 106 at the same time as the corresponding predicted MCPEL block of pels passes into DCT 126. The low frequency coefficients supplied as an output from DCT 126 appear as signal PRED, which is the aforementioned prediction for the low frequency coefficients LFREQ. DCT 126 does not generate any high frequency coefficients.
0015A corresponding example decoder is shown in FIG. 2. Channel decoder 202 decodes a signal supplied from the channel or from storage and supplies as outputs signals QHFREQ2, QLPERR2 and MOVECT2. Under errorless conditions, a signal with a suffix of 2 is a reconstructed version of the signal having the same base name as shown in Fig. 1. High frequency coefficients QHFREQ2 are inverse transformed by INVERSE DCT 204 and the resulting signal, HPEL, is supplied as an output. Signal HPEL represents a high frequency pel domain version of the original image. Quantized low frequency coefficient prediction error signal QLPERR2 is added to prediction signal PRED2 by adder 206 to generate reconstructed low frequency coefficients RLFREQ2, which are supplied as an output. Reconstructed low frequency coefficients RLFREQ2 are inverse transformed by INVERSE DCT 208 to form a signal LPEL2 which represents a low frequency motion compensated pel domain version of the original image. Adder 210 sums signals LPEL2 and HPEL and supplies the result as output signal VIDOUT. Signal VIDOUT is a full frequency reconstructed version of the original video image signal in the pel domain.
0016Additionally, low frequency pel signal LPEL2 is supplied to a frame memory 211 of motion compensation prediction unit 212. Motion compensation prediction unit 212 is also supplied with motion vectors MOVECT2. A shifted block of pels, signal MCPEL2, which represents an uncorrected motion compensation predicted version of the low frequency pel domain image is generated by motion compensation prediction unit 212. Signal MCPEL2 is supplied to DCT 214 wherein it is transformed into the frequency domain. The resulting output of DCT 214 is a set of low frequency prediction coefficients, PRED2, which are supplied to abovementioned adder 206. DCT 214 also does not generate any high frequency coefficients.
0017At start up, to initialize the motion compensation unit and establish a base frame, motion compensation of the low frequency image is suspended for one frame. Thereafter, at predetermined intervals, motion compensation of the low frequency image may be suspended for additional frame periods. This is useful in applications requiring random access from storage such as compact disks or where the image signal is to be transported over a packet network that cannot guaranty 100% delivery for packets that contain the motion vectors and low frequency error signal. Networks that guaranty 100% delivery would not require suspension of motion compensation.
0018In the encoder (FIG. 1), the time of suspension is determined by a control mechanism (not shown) and effected by changing the logical condition of signal CTRL supplied to motion compensation prediction unit 124. Suspension of motion compensation causes signals MOVECT, MCPEL and PRED to have values of zero (0). Nothing is therefore subtracted from the low frequency coefficients at subtracter 112 and LPERR and low frequency coefficients LFREQ are therefore equal. This causes the low frequency coefficients LFREQ to be directly encoded rather than prediction error signal LPERR and motion vectors MOVECT thereby resulting in the transmission of a new base frame.
0019Similarly in the receiver (FIG. 2), a corresponding time of suspension of motion compensation is determined by a control mechanism (also not shown) and effected by changing the logical condition of signal CTRL supplied to motion compensation prediction unit 212. When motion compensation is suspended in the encoder (FIG. 1) signal QLPERR2 in the receiver (FIG. 2) is equal to the low frequency coefficients LFREQ (FIG. 1). Suspension of motion compensation causes signals MCPEL2 and PRED2 to have values of zero (0). In turn, signal RLFREQ2 is equal to signal QLPERR2 and correspondingly LPEL2 represents a low frequency base image.
0020Separating the DCT coefficients has several advantages. First, the low frequency image and the high frequency image can be computed in parallel, since errors introduced by compressing the low frequency information do not translate into additional errors in the high frequency information. Also, since not all DCT coefficients are necessary to generate either the low frequency image or the high frequency image, the entire DCT need not be computed for each, and an intelligent implementation of the DCT can improve the efficiency of the computation. Additionally, a better prediction of the coefficients in the motion compensation loop is obtained, since fewer coefficients are predicted. Furthermore, by motion compensating only the low frequencies, the overall image quality is not sacrificed in areas with large motion. In this case, motion compensation cannot adequately predict the high frequencies anyway, so little is lost by not motion compensating them. This coding method has all the advantages for packet transmission inherent in any variable bit-rate method. A balance is struck between data compression and robustness to packet loss, by motion compensating low frequencies, yet intraframe coding the high frequencies.
0021An embedded code, as is well known, is a digital code that allows bits or symbols to be dropped and inserted without causing mismatch between the memories of the encoder and the decoder. The encoding and decoding system described above is such a system. Thus, another advantage of this system is that if some of the high frequency coefficients transmitted by the encoder are lost they can be zero-filled at the decoder and an image with less error than prior systems will be displayed. Further, only the low frequency portions of the intraframe coded base frame and subsequent interframe coded frames until ad including a desired random frame need be read. These are then combined with the high frequency portion of the desired frame and displayed. Therefore, a high quality image can be randomly accessed more rapidly than with prior coded video systems.
3 sheets
Sheet 1 Sheet 2 Sheet 3
Every citation, both ways
| Document | Relation | Office | Cited during |
|---|---|---|---|
| US6977961B1 | Cited by | United States of America | Applicant |
| FR2906433A1 | Cited by | France | Search report |
| WO9430010A1 | Cited by | World Intellectual Property Organization (WIPO) | International search |
| US8711945B2 | Cited by | United States of America | Applicant |
| US6977961B1 | Cited by | United States of America | Applicant |
| US6826232B2 | Cited by | United States of America | Applicant |
| WO0117266A1 | Cited by | World Intellectual Property Organization (WIPO) | International search |
| FR2906433A1 | Cited by | France | Search report |
| EP1081958A1 | Cited by | European Patent Office (EPO) | Search report |
| US5754246A | Cited by | United States of America | Search report |
| EP0339589A2 | Cites | European Patent Office (EPO) | Search report |
| US4245248A | Cites | United States of America | Search report |
8 members in 5 offices
Priority claims4
| Document | Office | Kind | Date |
|---|---|---|---|
| 517991 | United States of America | – | |
| 51799190 | United States of America | A | |
| 517991 | – | – | – |
| US19900517991 | – | – | – |
Members8
| Document | Office | Kind | |
|---|---|---|---|
| US5001561A | United States of America | A | |
| FI912110A0 | Finland | A0 | |
| CA2034418A1 | Canada | A1 | |
| FI912110A | Finland | A | |
| FI912110A7 | Finland | A7 | |
| EP0454927A2This record | European Patent Office (EPO) | A2 | |
| JPH04229791A | Japan | A | |
| EP0454927A3 | European Patent Office (EPO) | A3 |
9 legal events, as the office reported them to INPADOC
Over the term
Point at a mark for the eventEvents
| Event | Code | |
|---|---|---|
| Application deemed to be withdrawnWithdrawn18D | 18D | |
| Information on the status of an ep patent application or granted ep patentGrantedSTATUS: THE APPLICATION IS DEEMED TO BE WITHDRAWNSTAA | STAA | |
| Party data changed (applicant data changed or rights of an application transferred)RAP3 | RAP3 | |
| First examination report despatched17Q | 17Q | |
| Request for examination filed17P | 17P | |
| Designated contracting statesAK | AK | |
| Search report despatchedORIGINAL CODE: 0009013PUAL | PUAL | |
| Designated contracting statesAK | AK | |
| Public reference made under article 153(3) epc to a published international application that has entered the european phaseORIGINAL CODE: 0009012PUAI | PUAI |
Numbers
- Publication
- 0454927
- Publication, DOCDB
- 0454927
- Publication, EPODOC
- EP0454927
- Application
- 90313839
- Application, DOCDB
- 90313839
- Application, EPODOC
- EP19900313839
Titles6
- German
- Kodierungssystem für Videosignale.
- English
- A coding system for video signals.
- French
- Système de codage pour signaux vidéo.
- German
- Kodierungssystem für Videosignale
- English
- A coding system for video signals
- French
- Système de codage pour signaux vidéo
Classification
- CPC, 3
- H04N19/60
- H04N19/30
- H04N19/619
- IPC, 4
- H04N5 92
- H04N7 26
- H04N7 30
- H04N7 50
Designated states1
- Contracting states, 1
- Sweden