Audio coding system using characteristics of a decoded signal to adapt synthesized spectral components
19 claims: 10 independent, 9 dependent
- 1A method for processing encoded audio information, wherein the method comprises:receiving the encoded audio information and obtaining therefrom subband signals representing some but not all spectral content of an audio signal;examining the subband signals to obtain a characteristic of the audio signal, wherein the characteristic is tonality or temporal shape;generating synthesized spectral components that have the characteristic of the audio signal;integrating the synthesized spectral components with the subband signals to generate a set of modified subband signals;and generating the audio information by applying a synthesis filterbank to the set of modified subband signals.
- 9The method of any one of claims 1 through 8 that obtains the characteristics of the audio signal by examining components of one or more subband signals in a first portion of spectrum;and generates the synthesized spectral components by copying one or more components of the subband signals in the first portion of spectrum to a second portion of spectrum to form synthesized subband signals and modifying the copied components such that the synthesized subband signals have the characteristic of the audio signal.
- 11An apparatus for processing encoded audio information, wherein the apparatus comprises:an input terminal (21;76;77) adapted to receive the encoded audio information;memory (73, 74);and processing circuitry (72) coupled to the input terminal and the memory;wherein the processing circuitry is adapted to: receive (22) the encoded audio information and obtain (24) therefrom subband signals representing some but not all spectral content of an audio signal;examine (25) the subband signals to obtain a characteristic of the audio signal, wherein the characteristic is tonality or temporal shape;generate (26) synthesized spectral components that have the characteristic of the audio signal;integrate (27) the synthesized spectral components with the subband signals to generate a set of modified subband signals;and generate the audio information by applying a synthesis filterbank (28) to the set of modified subband signals.
- 19The apparatus of any one of claims 11 through 18, wherein the processing circuitry (72) is adapted to:obtain the characteristics of the audio signal by examining components of one or more subband signals in a first portion of spectrum;and generate the synthesized spectral components by copying one or more components of the subband signals in the first portion of spectrum to a second portion of spectrum to form synthesized subband signals and modifying the copied components such that the synthesized subband signals have the characteristic of the audio signal.
Independent claims11
67 paragraphs in 5 sections, as filed
TECHNICAL FIELD
0001The present invention is related generally to audio coding systems, and is related more specifically to improving the perceived quality of the audio signals obtained from audio coding systems.
BACKGROUND ART
0002Audio coding systems are used to encode an audio signal into an encoded signal that is suitable for transmission or storage, and then subsequently receive or retrieve the encoded signal and decode it to obtain a version of the original audio signal for playback. Perceptual audio coding systems attempt to encode an audio signal into an encoded signal that has lower information capacity requirements than the original audio signal, and then subsequently decode the encoded signal to provide an output that is perceptually indistinguishable from the original audio signal. One example of a perceptual audio coding system is described in the Advanced Television Systems Committee (ATSC) A/52A document entitled "Revision A to Digital Audio Compression (AC-3) Standard" published August 20, 2001, which is referred to as Dolby Digital. Another example is described in <nplcit id="ncit0001" npl-type="s"><text>Bosi et al., "ISO/IEC MPEG-2 Advanced Audio Coding." J. AES, vol. 45, no. 10, October 1997, pp. 789-814</text></nplcit>, which is referred to as Advanced Audio Coding (AAC). In these two coding systems, as well as in many other perceptual coding systems, a split-band transmitter applies an analysis filterbank to an audio signal to obtain spectral components that are arranged in groups or frequency bands, and encodes the spectral components according to psychoacoustic principles to generate an encoded signal. The band widths typically vary and are usually commensurate with widths of the so called critical bands of the human auditory system. A complementary split-band receiver receives and decodes the encoded signal to recover spectral components and applies a synthesis filterbank to the decoded spectral components to obtain a replica of the original audio signal.
0003Perceptual coding systems can be used to reduce the information capacity requirements of an audio signal while preserving a subjective or perceived measure of audio quality so that an encoded representation of the audio signal can be conveyed through a communication channel using less bandwidth or stored on a recording medium using less space. Information capacity requirements are reduced by quantizing the spectral components. Quantization injects noise into the quantized signal, but perceptual audio coding systems generally use psychoacoustic models in an attempt to control the amplitude of quantization noise so that it is masked or rendered inaudible by spectral components in the signal.
0004Traditional perceptual coding techniques work reasonably well in audio coding systems that are allowed to transmit or record encoded signals having medium to high bit rates, but these techniques by themselves do not provide very good audio quality when the encoded signals are constrained to low bit rates. Other techniques have been used in conjunction with perceptual coding techniques in an attempt to provide high quality signals at very low bit rates.
0005One technique called "High-Frequency Regeneration" (HFR) is described in <patcit id="pcit0001" dnum="US20030187663A1" dnum-type="L"><text>U.S. patent application publication number 2003-0187,663 A1</text></patcit>, entitled "Broadband Frequency Translation for High Frequency Regeneration" by Truman, et al., published October 2, 2003. In an audio coding system that uses HFR, a transmitter excludes high-frequency components from the encoded signal and a receiver regenerates or synthesizes noise-like substitute components for the missing high-frequency components. The resulting signal provided at the output of the receiver generally is not perceptually identical to the original signal provided at the input to the transmitter but sophisticated regeneration techniques can provide an output signal that is a fairly good approximation of the original input signal having a much higher perceived quality that would otherwise be possible at low bit rates. In this context, high quality usually means a wide bandwidth and a low level of perceived noise.
0006Another synthesis technique called "Spectral Hole Filling" (SHF) is described in <patcit id="pcit0002" dnum="US20030233234A1" dnum-type="L"><text>U.S. patent application publication number 2003-0233234 A1</text></patcit> entitled "Improved Audio Coding System Using Spectral Hole Filling" by Truman, et al., published December 18, 2003. According to this technique, a transmitter quantizes and encodes spectral components of an input signal in such a manner that bands of spectral components are omitted from the encoded signal. The bands of missing spectral components are referred to as spectral holes. A receiver synthesizes spectral components to fill the spectral holes. The SHF technique generally does not provide an output signal that is perceptually identical to the original input signal but it can improve the perceived quality of the output signal in systems that are constrained to operate with low bit rate encoded signals.
0007Another way of implementing said "Spectral Hole Filling" is disclosed in patent application <patcit id="pcit0003" dnum="WO0045379A"><text>WO00/45379</text></patcit>.
0008Techniques like HFR and SHF can provide an advantage in many situations but they do not work well in all situations. One situation that is particularly troublesome arises when an audio signal having a rapidly changing amplitude is encoded by a system that uses block transforms to implement the analysis and synthesis filterbanks. In this situation, audible noise-like components can be smeared across a period of time that corresponds to a transform block.
0009One technique that can be used to reduce the audible effects of time-smeared noise is to decrease the block length of the analysis and synthesis transforms for intervals of the input signal that are highly non-stationary. This technique works well in audio coding systems that are allowed to transmit or record encoded signals having medium to high bit rates, but it does not work as well in lower bit rate systems because the use of shorter blocks reduces the coding gain achieved by the transform.
0010In another technique, a transmitter modifies the input signal so that rapid changes in amplitude are removed or reduced prior to application of the analysis transform. The receiver reverses the effects of the modifications after application of the synthesis transform. Unfortunately, this technique obscures the true spectral characteristics of the input signal, thereby distorting information needed for effective perceptual coding, and because the transmitter must use part of the transmitted signal to convey parameters that the receiver needs to reverse the effects of the modifications.
0011In a third technique known as temporal noise shaping, a transmitter applies a prediction filter to the spectral components obtained from the analysis filterbank, conveys prediction errors and the predictive filter coefficients in the transmitted signal, and the receiver applies an inverse prediction filter to the prediction errors to recover the spectral components. This technique is undesirable in low bit rate systems because of the signal overhead needed to convey the predictive filter coefficients.
DISCLOSURE OF INVENTION
0012It is an object of the present invention to provide techniques that can be used in low bit rate audio coding systems to improve the perceived quality of the audio signals generated by such systems.
0013According to the present invention, encoded audio information is processed by receiving the encoded audio information and obtaining subband signals representing some but not all spectral content of an audio signal, examining the subband signals to obtain a characteristic of the audio signal, where the characteristic is tonality or temporal shape, generating synthesized spectral components that have the characteristic of the audio signal, integrating the synthesized spectral components with the subband signals to generate a set of modified subband signals, and generating the audio information by applying a synthesis filterbank to the set of modified subband signals.
0014The various features of the present invention and its preferred embodiments may be better understood by referring to the following discussion and the accompanying drawings. The contents of the following discussion and the drawings are set forth as examples only and should not be understood to represent limitations upon the scope of the present invention.
BRIEF DESCRIPTION OF DRAWINGS
0015<ul id="ul0001" list-style="none" compact="compact"><li><figref idref="f0001">Fig. 1</figref> is a schematic block diagram of a transmitter in an audio coding system.</li><li><figref idref="f0001">Fig. 2</figref> is a schematic block diagram of a receiver in an audio coding system.</li><li><figref idref="f0002">Fig. 3</figref> is a schematic block diagram of an apparatus that may be used to implement various aspects of the present invention.</li></ul>
MODES FOR CARRYING OUT THE INVENTION
A. Overview
0016Various aspects of the present invention may be incorporated into a variety of signal processing methods and devices including devices like those illustrated in <figref idref="f0001">Figs. 1and 2</figref>. Some aspects may be carried out by processing performed in only a receiver. Other aspects require cooperative processing performed in both a receiver and a transmitter. A description of processes that may be used to carry out these various aspects of the present invention is provided below following an overview of typical devices that may be used to perform these processes.
0017<figref idref="f0001">Fig 1</figref> illustrates one implementation of a split-band audio transmitter in which the analysis filterbank 12 receives from the path 11 audio information representing an audio signal and, in response, provides frequency subband signals that represent spectral content of the audio signal. Each subband signal is passed to the encoder 14, which generates an encoded representation of the subband signals and passes the encoded representation to the formatter 16. The formatter 16 assembles the encoded representation into an output signal suitable for transmission or storage, and passes the output signal along the path 17.
0018<figref idref="f0001">Fig 2</figref> illustrates one implementation of a split-band audio receiver in which the deformatter 22 receives from the path 21 an input signal conveying an encoded representation of frequency subband signals representing spectral content of an audio signal. The deformatter 22 obtains the encoded representation from the input signal and passes it to the decoder 24. The decoder 24 decodes the encoded representation into frequency subband signals. The analyzer 25 examines the subband signals to obtain one or more characteristics of the audio signal that the subband signals represent. An indication of the characteristics is passed to the component synthesizer 26, which generates synthesized spectral components using a process that adapts in response to the characteristics. The integrator 27 generates a set of modified subband signals by integrating the subband signals provided by the decoder 24 with the synthesized spectral components generated by the component synthesizer 26. In response to the set of modified subband signals, the synthesis filterbank 28 generates along the path 29 audio information representing an audio signal. In the particular implementation shown in the figure, neither the analyzer 25 nor the component synthesizer 26 adapt processing in response to any control information obtained from the input signal by the deformatter 22. In other implementations, the analyzer 25 and/or the component synthesizer 26 can be responsive to control information obtained from the input signal.
0019The devices illustrated in <figref idref="f0001">Figs. 1 and 2</figref> show filterbanks for three frequency subbands. Many more subbands are used in a typical implementation but only three are shown for illustrative clarity. No particular number is important to the present invention.
0020The analysis and synthesis filterbanks may be implemented by essentially any block transform including a Discrete Fourier Transform or a Discrete Cosine
0021Transform (DCT). In one audio coding system having a transmitter and a receiver like those discussed above, the analysis filterbank 12 and the synthesis filterbank 28 are implemented by modified DCT known as Time-Domain Aliasing Cancellation (TDAC) transforms, which are described in <nplcit id="ncit0002" npl-type="s"><text>Princen et al., "Subband/Transform Coding Using Filter Bank Designs Based on Time Domain Aliasing Cancellation," ICASSP 1987 Conf. Proc., May 1987, pp. 2161-64</text></nplcit>.
0022Analysis filterbanks that are implemented by block transforms convert a block or interval of an input signal into a set of transform coefficients that represent the spectral content of that interval of signal. A group of one or more adjacent transform coefficients represents the spectral content within a particular frequency subband having a bandwidth commensurate with the number of coefficients in the group. The term "subband signal" refers to groups of one or more adjacent transform coefficients and the term "spectral components" refers to the transform coefficients.
0023The terms "encoder" and "encoding" used in this disclosure refer to information processing devices and methods that may be used to represent an audio signal with encoded information having lower information capacity requirements than the audio signal itself. The terms "decoder" and "decoding" refer to information processing devices and methods that may be used to recover an audio signal from the encoded representation. Two examples that pertain to reduced information capacity requirements are the coding needed to process bit streams compatible with the Dolby Digital and the AAC coding standards mentioned above. No particular type of encoding or decoding is important to the present invention.
B. Receiver
0024Various aspects of the present invention may be carried out in a receiver that do not require any special processing or information from a transmitter. These aspects are described first.
1. Analysis of Signal Characteristics
0025The present invention may be used in coding systems that represent audio signals with very low bit rate encoded signals. The encoded information in very low bit rate systems typically conveys subband signals that represent only a portion of the spectral components of the audio signal. The analyzer 25 examines these subband signals to obtain one or more characteristics of tonality and temporal shape of the portion of the audio signal that is represented by the subband signals. Representations of the one or more characteristics are passed to the component synthesizer 26 and are used to adapt the generation of synthesized spectral components. Several examples of characteristics in addition to tonality and temporal shape that may also be used are described below.
<i>a) Amplitude</i>
0026The encoded information generated by many coding systems represents spectral components that have been quantized to some desired bit length or quantizing resolution. Small spectral components having magnitudes less than the level represented by the least-significant bit (LSB) of the quantized components can be omitted from the encoded information or, alternatively, represented in some form that indicates the quantized value is zero or deemed to be zero. The level corresponding to the LSB of the quantized spectral components that are conveyed by the encoded information can be considered an upper bound on the magnitude of the small spectral components that are omitted from the encoded information.
0027The component synthesizer 26 can use this level to limit the amplitude of any component that is synthesized to replace a missing spectral component.
<i>b) Spectral Shape</i>
0028The spectral shape of the subband signals conveyed by the encoded information is immediately available from the subband signals themselves; however, other information about spectral shape can be derived by applying a filter to the subband signals in the frequency domain. The filter may be a prediction filter, a lowpass filter, or essentially any other type of filter that may be desired.
0029An indication of the spectral shape or the filter output is passed to the component synthesizer 26 as appropriate. If necessary, an indication of which filter is used should also be passed.
<i>c) Masking</i>
0030A perceptual model may be applied to estimate the psychoacoustic masking effects of the spectral components in the subband signals. Because these masking effects vary by frequency, the masking provided by a first spectral component at one frequency will not necessarily provide the same level of masking as that provided by a second spectral component at another frequency even though the first and second spectral component have the same amplitude.
0031An indication of estimated masking effects is passed to the component synthesizer 26, which controls the synthesis of spectral components so that the estimated masking effects of the synthesized components have a desired relationship with the estimated masking effects of the spectral components in the subband signals.
<i>d) Tonality</i>
0032The tonality of the subband signals can be assessed in a variety of ways including the calculation of a Spectral Flatness Measure, which is a normalized quotient of the arithmetic mean of subband signal samples divided by the geometric mean of the subband signal samples. Tonality can also be assessed by analyzing the arrangement or distribution of spectral components within the subband signals. For example, a subband signal may be deemed to be more tonal rather than more like noise if a few large spectral components are separated by long intervals of much smaller components. Yet another way applies a prediction filter to the subband signals to determine the prediction gain. A large prediction gain tends to indicate a signal is more tonal.
0033An indication of tonality is passed to the component synthesizer 26, which controls synthesis so that the synthesized spectral component have an appropriate level of tonality. This may be done by forming a weighted combination of tone-like and noise-like synthesized components to achieve the desired level of tonality.
<i>e) Temporal Shape</i>
0034The temporal shape of a signal represented by subband signals can be estimated directly from the subband signals. The technical basis for one implementation of a temporal-shape estimator may be explained in terms of a linear system represented by equation 1. <maths id="math0001" num="(1)"><math display="block"><mi>y</mi><mfenced><mi>t</mi></mfenced><mo>=</mo><mi>h</mi><mfenced><mi>t</mi></mfenced><mo>⋅</mo><mi>x</mi><mfenced><mi>t</mi></mfenced></math><img file="EP1514263B1_D0001.tif" /></maths> where <ul id="ul0002" list-style="none" compact="compact"><li><i>y</i>(<i>t</i>) = a signal having a temporal shape to be estimated;</li><li><i>h</i>(<i>t</i>) = the temporal shape of the signal <i>y</i>(<i>t</i>);</li><li>the dot symbol (·) denotes multiplication; and</li><li><i>x</i>(<i>t</i>) = a temporally-flat version of the signal <i>y</i>(<i>t</i>).</li></ul>
0035This equation may be rewritten as: <maths id="math0002" num="(2)"><math display="block"><mi>Y</mi><mfenced open="[" close="]"><mi>k</mi></mfenced><mo>=</mo><mi>H</mi><mfenced open="[" close="]"><mi>k</mi></mfenced><mo>*</mo><mi>X</mi><mfenced open="[" close="]"><mi>k</mi></mfenced></math><img file="EP1514263B1_D0002.tif" /></maths> where <ul id="ul0003" list-style="none" compact="compact"><li><i>Y</i>[<i>k</i>] = a frequency-domain representation of the signal <i>y</i>(<i>t</i>);</li><li><i>H</i>[<i>k</i>] = a frequency-domain representation of <i>h</i>(<i>t</i>);</li><li>the star symbol (*) denotes convolution; and</li><li><i>X</i>[<i>k</i>] = a frequency-domain representation of the signal <i>x</i>(<i>t</i>).</li></ul>
0036The frequency-domain representation <i>Y</i>[<i>k</i>] corresponds to one or more of the subband signals obtained by the decoder 24. The analyzer 25 can obtain an estimate of the frequency-domain representation <i>H</i>[<i>k</i>] of the temporal shape <i>h</i>(<i>t</i>) by solving a set of equations derived from an autoregressive moving average (ARMA) model of <i>Y</i>[<i>k</i>] and <i>X</i>[<i>k</i>]. Additional information about the use of ARMA models may be obtained from <nplcit id="ncit0003" npl-type="b"><text>Proakis and Manolakis, "Digital Signal Processing: Principles, Algorithms and Applications," MacMillan Publishing Co., New York, 1988</text></nplcit>. See especially pp. 818-821.
0037The frequency-domain representation <i>Y</i>[<i>k</i>] is arranged in blocks of transform coefficients. Each block of transform coefficients expresses a short-time spectrum of the signal <i>y</i>(<i>t</i>). The frequency-domain representation <i>X</i>[<i>k</i>] is also arranged in blocks. Each block of coefficients in the frequency-domain representation <i>X</i>[<i>k</i>] represents a block of samples for the temporally-flat signal <i>x</i>(<i>t</i>) that is assumed to be wide sense stationary. It is also assumed the coefficients in each block of the <i>X</i>[<i>k</i>] representation are independently distributed. Given these assumptions, the signals can be expressed by an ARMA model as follows: <maths id="math0003" num="(3)"><math display="block"><mi>Y</mi><mfenced open="[" close="]"><mi>k</mi></mfenced><mo>+</mo><mstyle displaystyle="true"><munderover><mo>∑</mo><mrow><mi>i</mi><mo>=</mo><mn>1</mn></mrow><mi>L</mi></munderover></mstyle><msub><mi>a</mi><mi>i</mi></msub><mo></mo><mi>Y</mi><mo></mo><mfenced open="[" close="]"><mi>k</mi><mo>-</mo><mi>l</mi></mfenced><mo>=</mo><mstyle displaystyle="true"><munderover><mo>∑</mo><mrow><mi>q</mi><mo>=</mo><mn>0</mn></mrow><mi>Q</mi></munderover></mstyle><msub><mi>b</mi><mi>q</mi></msub><mo></mo><mi>X</mi><mo></mo><mfenced open="[" close="]"><mi>k</mi><mo>-</mo><mi>q</mi></mfenced></math><img file="EP1514263B1_D0003.tif" /></maths> where <ul id="ul0004" list-style="none" compact="compact"><li><i>L</i> = length of the autoregessive portion of the ARMA model; and</li><li><i>Q</i> = the length of the moving average portion of the ARMA model.</li></ul>
0038Equation 3 can be solved for <i>a<sub>l</sub></i> and <i>b<sub>q</sub></i> by solving for the autocorrelation of <i>Y[k]:</i><maths id="math0004" num="(4)"><math display="block"><mi>E</mi><mfenced open="{" close="}"><mi>Y</mi><mfenced open="[" close="]"><mi>k</mi></mfenced><mo>⋅</mo><mi>Y</mi><mo></mo><mfenced open="[" close="]"><mi>k</mi><mo>-</mo><mi>m</mi></mfenced></mfenced><mo>=</mo><mo>-</mo><mstyle displaystyle="true"><munderover><mo>∑</mo><mrow><mi>l</mi><mo>=</mo><mn>1</mn></mrow><mi>L</mi></munderover></mstyle><msub><mi>a</mi><mi>l</mi></msub><mo></mo><mi>E</mi><mfenced open="{" close="}"><mi>Y</mi><mo></mo><mfenced open="[" close="]"><mi>k</mi><mo>-</mo><mi>l</mi></mfenced><mo>⋅</mo><mi>Y</mi><mo></mo><mfenced open="[" close="]"><mi>k</mi><mo>-</mo><mi>m</mi></mfenced></mfenced><mo>+</mo><mstyle displaystyle="true"><munderover><mo>∑</mo><mrow><mi>q</mi><mo>=</mo><mn>0</mn></mrow><mi>Q</mi></munderover></mstyle><msub><mi>b</mi><mi>q</mi></msub><mo></mo><mi>E</mi><mfenced open="{" close="}"><mi>X</mi><mo></mo><mfenced open="[" close="]"><mi>k</mi><mo>-</mo><mi>q</mi></mfenced><mo>⋅</mo><mi>Y</mi><mo></mo><mfenced open="[" close="]"><mi>k</mi><mo>-</mo><mi>m</mi></mfenced></mfenced></math><img file="EP1514263B1_D0004.tif" /></maths> where <i>E</i>{} denotes the expected value function.
0039Equation 4 can be rewritten as: <maths id="math0005" num="(5)"><math display="block"><msub><mi>R</mi><mi mathvariant="italic">YY</mi></msub><mfenced open="[" close="]"><mi>m</mi></mfenced><mo>=</mo><mo>-</mo><mstyle displaystyle="true"><munderover><mo>∑</mo><mrow><mi>l</mi><mo>=</mo><mn>1</mn></mrow><mi>L</mi></munderover></mstyle><msub><mi>a</mi><mi>l</mi></msub><mo></mo><msub><mi>R</mi><mi mathvariant="italic">YY</mi></msub><mo></mo><mfenced><mi>m</mi><mo>-</mo><mi>l</mi></mfenced><mo>+</mo><mstyle displaystyle="true"><munderover><mo>∑</mo><mrow><mi>q</mi><mo>=</mo><mn>0</mn></mrow><mi>Q</mi></munderover></mstyle><msub><mi>b</mi><mi>q</mi></msub><mo></mo><msub><mi>R</mi><mrow><mi mathvariant="italic">X</mi><mo></mo><mi mathvariant="italic">Y</mi></mrow></msub><mo></mo><mfenced open="[" close="]"><mi>m</mi><mo>-</mo><mi>q</mi></mfenced></math><img file="EP1514263B1_D0005.tif" /></maths> where <ul id="ul0005" list-style="none" compact="compact"><li><i>R<sub>YY</sub></i>[<i>n</i>] denotes the autocorrelation of <i>Y</i>[<i>n</i>]; and</li><li><i>R<sub>XY</sub></i>[<i>k</i>] denotes the cross-correlation of <i>Y</i>[<i>k</i>] and <i>X</i>[<i>k</i>].</li></ul>
0040If we further assume the linear system represented by <i>H</i>[<i>k</i>] is only autoregressive, then the second term on the right side of equation 5 can be ignored. Equation 5 can then be rewritten as: <maths id="math0006" num="(6)"><math display="block"><msub><mi>R</mi><mi mathvariant="italic">YY</mi></msub><mfenced open="[" close="]"><mi>m</mi></mfenced><mo>=</mo><mo>-</mo><mstyle displaystyle="true"><munderover><mo>∑</mo><mrow><mi>i</mi><mo>=</mo><mn>1</mn></mrow><mi>L</mi></munderover></mstyle><msub><mi>a</mi><mi>l</mi></msub><mo></mo><msub><mi>R</mi><mi mathvariant="italic">YY</mi></msub><mo></mo><mfenced><mi>m</mi><mo>-</mo><mi>l</mi></mfenced><mspace width="1em" /><mi>for</mi><mspace width="1em" /><mi>m</mi><mo>></mo><mn>0</mn></math><img file="EP1514263B1_D0006.tif" /></maths> which represents a set of <i>L</i> linear equations that can be solved to obtain the the <i>L</i> coefficients <i>a<sub>i</sub></i>.
0041With this explanation, it is now possible to describe one implementation of a temporal-shape estimator that uses frequency-domain techniques. In this implementation, the temporal-shape estimator receives the frequency-domain representation <i>Y</i>[<i>k</i>] of one or more subband signals <i>y</i>(<i>t</i>) and calculates the autocorrelation sequence <i>R<sub>YY</sub></i>[<i>m</i>] for -<i>L</i> ≤ <i>m</i> ≤ <i>L.</i> These values are used to establish a set of linear equations that are solved to obtain the coefficients <i>a<sub>i</sub></i>, which represent the poles of a linear all-pole filter <i>FR</i> shown below in equation 7. <maths id="math0007" num="(7)"><math display="block"><mi mathvariant="italic">FR</mi><mfenced><mi>z</mi></mfenced><mo>=</mo><mfrac><mn>1</mn><mrow><mn>1</mn><mo>+</mo><mstyle displaystyle="true"><munderover><mo>∑</mo><mrow><mi>i</mi><mo>=</mo><mn>1</mn></mrow><mi>L</mi></munderover><msub><mi>a</mi><mi>i</mi></msub></mstyle><msup><mi>z</mi><mrow><mo>-</mo><mi>i</mi></mrow></msup></mrow></mfrac></math><img file="EP1514263B1_D0007.tif" /></maths> This filter can be applied to the frequency-domain representation of an arbitrary temporally-flat signal such as a noise-like signal to obtain a frequency-domain representation of a version of that temporally-flat signal having a temporal shape substantially equal to the temporal shape of the signal <i>y</i>(<i>t</i>).
0042A description of the poles of filter <i>FR</i> may be passed to the component synthesizer 26, which can use the filter to generate synthesized spectral components representing a signal having the desired temporal shape.
2. Generation of Synthesized Components
0043The component synthesizer 26 may generate the synthesized spectral components in a variety of ways. Two ways are described below. Multiple ways may be used. For example, different ways may be selected in response to characteristics derived from the subband signals or as a function of frequency.
0044A first way generates a noise-like signal. For example, essentially any of a wide variety of time-domain and frequency-domain techniques may be used to generate noise-like signals.
0045A second way uses a frequency-domain technique called spectral translation or spectral replication that copies spectral components from one or more frequency subbands. Lower-frequency spectral components are usually copied to higher frequencies because higher frequency components are often related in some manner to lower frequency components. In principle, however, spectral components may be copied to higher or lower frequencies. If desired, noise may be added or blended with the translated components and the amplitude may be modified as desired. Preferably, adjustments are made as necessary to eliminate or at least reduce discontinuities in the phase of the synthesized components.
0046The synthesis of spectral components is controlled by information received from the analyzer 25 so that the synthesized components have one or more characteristics obtained from the subband signals.
3. Integration of Signal Components
0047The synthesized spectral components may be integrated with the subband signal spectral components in a variety of ways. One way uses the synthesized components as a form of dither by combining respective synthesized and subband components representing corresponding frequencies. Another way substitutes one or more synthesized components for selected spectral components that are present in the subband signals. Yet another way merges synthesized components with components of the subband signals to represent spectral components that are not present in the subband signals. These and other ways may be used in various combinations.
C. Transmitter
0048Aspects of the present invention described above can be carried out in a receiver without requiring the transmitter to provide any control information beyond what is needed by a receiver to receive and decode the subband signals without features of the present invention. These aspects of the present invention can be enhanced if additional control information is provided. One example is discussed below.
0049The degree to which temporal shaping is applied to the synthesized components can be adapted by control information provided in the encoded information. One way this can be done is through the use of a parameter β as shown in the following equation. <maths id="math0008" num="(8)"><math display="block"><mi mathvariant="italic">FR</mi><mfenced><mi>z</mi></mfenced><mo>=</mo><mfrac><mn>1</mn><mrow><mn>1</mn><mo>+</mo><mstyle displaystyle="true"><munderover><mo>∑</mo><mrow><mi>i</mi><mo>=</mo><mn>1</mn></mrow><mi>L</mi></munderover><msub><mi>a</mi><mi>i</mi></msub></mstyle><msup><mi>β</mi><mi>i</mi></msup><mo></mo><msup><mi>z</mi><mrow><mo>-</mo><mi>i</mi></mrow></msup></mrow></mfrac><mspace width="1em" /><mi>for</mi><mspace width="1em" /><mn>0</mn><mo>≤</mo><mi mathvariant="normal">β</mi><mo>≤</mo><mn>1</mn></math><img file="EP1514263B1_D0008.tif" /></maths> The filter provides no temporal shaping when β=0. When β=1, the filter provides a degree of temporal shaping such that correlation between the temporal shape of the synthesized components and the temporal shape of the subband signals is maximum. Other values for β provide intermediate levels of temporal shaping.
0050In one implementation, the transmitter provides control information that allows the receiver to set β to one of eight values.
0051The transmitter may provide other control information that the receiver can use to adapt the component synthesis process in any way that may be desired.
D. Implementation
0052Various aspects of the present invention may be implemented in a wide variety of ways including software in a general-purpose computer system or in some other apparatus that includes more specialized components such as digital signal processor (DSP) circuitry coupled to components similar to those found in a general-purpose computer system. <figref idref="f0002">Fig. 3</figref> is a block diagram of device 70 that may be used to implement various aspects of the present invention in transmitter or receiver. DSP 72 provides computing resources. RAM 73 is system random access memory (RAM) used by DSP 72 for signal processing. ROM 74 represents some form of persistent storage such as read only memory (ROM) for storing programs needed to operate device 70 and to carry out various aspects of the present invention. I/O control 75 represents interface circuitry to receive and transmit signals by way of communication channels 76, 77. Analog-to-digital converters and digital-to-analog converters may be included in I/O control 75 as desired to receive and/or transmit analog audio signals. In the embodiment shown, all major system components connect to bus 71, which may represent more than one physical bus; however, a bus architecture is not required to implement the present invention.
0053In embodiments implemented in a general purpose computer system, additional components may be included for interfacing to devices such as a keyboard or mouse and a display, and for controlling a storage device having a storage medium such as magnetic tape or disk, or an optical medium. The storage medium may be used to record programs of instructions for operating systems, utilities and applications, and may include embodiments of programs that implement various aspects of the present invention.
0054The functions required to practice various aspects of the present invention can be performed by components that are implemented in a wide variety of ways including discrete logic components, one or more ASICs and/or program-controlled processors. The manner in which these components are implemented is not important to the present invention.
0055Software implementations of the present invention may be conveyed by a variety machine readable media such as baseband or modulated communication paths throughout the spectrum including from supersonic to ultraviolet frequencies, or storage media including those that convey information using essentially any magnetic or optical recording technology including magnetic tape, magnetic disk, and optical disc. Various aspects can also be implemented in various components of computer system 70 by processing circuitry such as ASICs, general-purpose integrated circuits, microprocessors controlled by programs embodied in various forms of ROM or RAM, and other techniques.
Contents5
10 sheets
Sheet 1 Sheet 2 Sheet 3 Sheet 4 Sheet 5 Sheet 6 Sheet 7 Sheet 8 Sheet 9 Sheet 10
Every citation, both ways
| Document | Relation | Office |
|---|---|---|
| WO0045379A | Cites | World Intellectual Property Organization (WIPO) |
| ATKINSON I A ET AL: "TIME ENVELOPE LP VOCODER: A NEW CODING TECHNIQUE AT VERY LOW BIT RATES" 4TH EUROPEAN CONFERENCE ON SPEECH COMMUNICATION AND TECHNOLOGY. EUROSPEECH '95. MADRID, SPAIN, SEPT. 18 - 21, 1995, EUROPEAN CONFERENCE ON SPEECH COMMUNICATION AND TECHNOLOGY. (EUROSPEECH), MADRID: GRAFICAS BRENS, ES, vol. 1 CONF. 4, 18 September 1995 (1995-09-18), pages 241-244, XP000854697 | Non-patent | – |
135 members in 21 offices
Priority claims14
| Document | Office | Kind | Date |
|---|---|---|---|
| 174493 | United States of America | – | |
| 17449302 | United States of America | A | |
| 17449302 | United States of America | A | |
| 238047 | United States of America | – | |
| 23804702 | United States of America | A | |
| 23804702 | United States of America | A | |
| 0318065 | United States of America | W | |
| 0318065 | United States of America | W | |
| 174493 | – | – | – |
| 238047 | – | – | – |
| US20020174493 | – | – | – |
| US20020238047 | – | – | – |
| US2003018065 | – | – | – |
| WO2003US18065 | – | – | – |
Members135
| Document | Office | Kind | |
|---|---|---|---|
| US2003233234A1 | United States of America | A1 | |
| US2003233236A1 | United States of America | A1 | |
| CA2489441A1 | Canada | A1 | |
| CA2489443A1 | Canada | A1 | |
| CA2735830A1 | Canada | A1 | |
| CA2736046A1 | Canada | A1 | |
| CA2736055A1 | Canada | A1 | |
| CA2736060A1 | Canada | A1 | |
| CA2736065A1 | Canada | A1 | |
| WO03107328A1 | World Intellectual Property Organization (WIPO) | A1 | |
| WO03107329A1 | World Intellectual Property Organization (WIPO) | A1 | |
| AU2003237295A1 | Australia | A1 | |
| AU2003243441A1 | Australia | A1 | |
| TW200400487A | Taiwan Province of China | A | |
| TW200404273A | Taiwan Province of China | A | |
| KR20050010945A | Republic of Korea | A | |
| KR20050010950A | Republic of Korea | A | |
| EP1514261A1 | European Patent Office (EPO) | A1 | |
| EP1514263A1 | European Patent Office (EPO) | A1 | |
| MXPA04012539A | Mexico | A | |
| MXPA04012540A | Mexico | A | |
| HK1070728A | Hong Kong, China | A | |
| HK1070728A1 | Hong Kong, China | A1 | |
| HK1070729A | Hong Kong, China | A | |
| HK1070729A1 | Hong Kong, China | A1 | |
| PL371898A1 | Poland | A1 | |
| PL372104A1 | Poland | A1 | |
| CN1662958A | China | A | |
| CN1662960A | China | A | |
| JP2005530205A | Japan | A | |
| JP2005530206A | Japan | A | |
| IL165648A0 | Israel | A0 | |
| IL165648D0 | Israel | D0 | |
| IL165650A0 | Israel | A0 | |
| IL165650D0 | Israel | D0 | |
| EP1514261B1 | European Patent Office (EPO) | B1 | |
| EP1736966A2 | European Patent Office (EPO) | A2 | |
| AT349754T | Austria | T | |
| ATE349754T1 | Austria | T1 | |
| DE60310716D1 | Germany | D1 | |
| DK1514261T3 | Denmark | T3 | |
| CN1310210C | China | C | |
| ES2275098T3 | Spain | T3 | |
| DE60310716T2 | Germany | T2 | |
| TWI288915B | Taiwan Province of China | B | |
| EP1736966A3 | European Patent Office (EPO) | A3 | |
| DE60310716T8 | Germany | T8 | |
| CN100369109C | China | C | |
| US7337118B2 | United States of America | B2 | |
| US2008140405A1 | United States of America | A1 | |
| MY136521A | Malaysia | A | |
| US7447631B2 | United States of America | B2 | |
| AU2003243441B2 | Australia | B2 | |
| US2009138267A1 | United States of America | A1 | |
| US2009144055A1 | United States of America | A1 | |
| AU2003243441C1 | Australia | C1 | |
| EP1514263B1This record | European Patent Office (EPO) | B1 | |
| KR20100063141A | Republic of Korea | A | |
| AT470220T | Austria | T | |
| ATE470220T1 | Austria | T1 | |
| JP4486496B2 | Japan | B2 | |
| EP1736966B1 | European Patent Office (EPO) | B1 | |
| EP2207169A1 | European Patent Office (EPO) | A1 | |
| EP2207170A1 | European Patent Office (EPO) | A1 | |
| AT473503T | Austria | T | |
| ATE473503T1 | Austria | T1 | |
| DE60332833D1 | Germany | D1 | |
| JP2010156990A | Japan | A | |
| EP2209115A1 | European Patent Office (EPO) | A1 | |
| KR20100086067A | Republic of Korea | A | |
| KR20100086068A | Republic of Korea | A | |
| EP2216777A1 | European Patent Office (EPO) | A1 | |
| DE60333316D1 | Germany | D1 | |
| KR100986150B1 | Republic of Korea | B1 | |
| KR100986152B1 | Republic of Korea | B1 | |
| KR100986153B1 | Republic of Korea | B1 | |
| DK1736966T3 | Denmark | T3 | |
| KR100991448B1 | Republic of Korea | B1 | |
| KR100991450B1 | Republic of Korea | B1 | |
| HK1141623A | Hong Kong, China | A | |
| HK1141623A1 | Hong Kong, China | A1 | |
| HK1141624A | Hong Kong, China | A | |
| HK1141624A1 | Hong Kong, China | A1 | |
| IL165650A | Israel | A | |
| PL207861B1 | Poland | B1 | |
| PL208344B1 | Poland | B1 | |
| HK1146145A | Hong Kong, China | A | |
| HK1146145A1 | Hong Kong, China | A1 | |
| HK1146146A | Hong Kong, China | A | |
| HK1146146A1 | Hong Kong, China | A1 | |
| EP2209115B1 | European Patent Office (EPO) | B1 | |
| US8032387B2 | United States of America | B2 | |
| AT526661T | Austria | T | |
| ATE526661T1 | Austria | T1 | |
| EP2207169B1 | European Patent Office (EPO) | B1 | |
| EP2207170B1 | European Patent Office (EPO) | B1 | |
| US8050933B2 | United States of America | B2 | |
| AT529858T | Austria | T | |
| AT529859T | Austria | T | |
| ATE529858T1 | Austria | T1 |
61 legal events, as 9 offices reported them to INPADOC
Over the term
Point at a mark for the eventEvents
| Event | Code | Office | |
|---|---|---|---|
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Ep patent has lapsedLapsedEUG | EUG | SE | |
| Patent expired after termination of 20 yearsExpiredPE20 | PE20 | GB | |
| Opt-out of the competence of the unified patent court (upc) registeredP01 | P01 | EP | |
| Patent ceasedCeasedPL | PL | CH | |
| Patent expired because of reaching the maximum lifetime of a patentExpiredMK | MK | NL | |
| Expiry of rightR071 | R071 | DE | |
| Annual fee paid to national office [announced via postgrant information from national office to epo]GrantedPGFP | PGFP | EP | |
| Annual fee paid to national office [announced via postgrant information from national office to epo]GrantedPGFP | PGFP | EP | |
| Annual fee paid to national office [announced via postgrant information from national office to epo]GrantedPGFP | PGFP | EP | |
| Annual fee paid to national office [announced via postgrant information from national office to epo]GrantedPGFP | PGFP | EP | |
| Annual fee paid to national office [announced via postgrant information from national office to epo]GrantedPGFP | PGFP | EP | |
| Annual fee paid to national office [announced via postgrant information from national office to epo]GrantedPGFP | PGFP | EP | |
| Fee paymentPLFP | PLFP | FR | |
| Fee paymentPLFP | PLFP | FR | |
| Fee paymentPLFP | PLFP | FR | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| No opposition filed against granted patent, or epo opposition proceedings concluded without decisionGrantedR097 | R097 | DE | |
| No opposition filed against granted patent, or epo opposition proceedings concluded without decisionGrantedR097 | R097 | DE | |
| No opposition filedOpposition26N | 26N | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| No opposition filed within time limitOppositionORIGINAL CODE: 0009261PLBE | PLBE | EP | |
| Information on the status of an ep patent application or granted ep patentGrantedSTATUS: NO OPPOSITION FILED WITHIN TIME LIMITSTAA | STAA | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Standard patents granted in hong kongGrantedGR | GR | HK | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Lapsed in a contracting state [announced via postgrant information from national office to epo]LapsedPG25 | PG25 | EP | |
| Translation of granted ep patentGrantedTRGR | TRGR | SE | |
| New agentNV | NV | CH | |
| Dpma publication of mentioned ep patent grantGrantedR096 | R096 | DE | |
| Corresponds to:REF | REF | EP | |
| European patents granted designating irelandGrantedFG4D | FG4D | IE | |
| Translation filed for an european patent granted for nl, confirming art. 52 par. 1 or 6 of the patents act 1995GrantedT3 | T3 | NL | |
| European patent takes effect as a national patent in ch/liEP | EP | CH | |
| Designated contracting statesAK | AK | EP | |
| European patent grantedGrantedFG4D | FG4D | GB | |
| (expected) grantORIGINAL CODE: 0009210GRAA | GRAA | EP | |
| Grant fee paidORIGINAL CODE: EPIDOSNIGR3GRAS | GRAS | EP | |
| Despatch of communication of intention to grant a patentORIGINAL CODE: EPIDOSNIGR1GRAP | GRAP | EP | |
| Request for extension of the european patent (deleted)DAX | DAX | EP | |
| Requests to designate patent in hong kongDE | DE | HK | |
| Request for examination filed17P | 17P | EP | |
| Designated contracting statesAK | AK | EP | |
| Request for extension of the european patentAX | AX | EP | |
| Public reference made under article 153(3) epc to a published international application that has entered the european phaseORIGINAL CODE: 0009012PUAI | PUAI | EP |
Numbers
- Publication
- 1514263
- Publication, DOCDB
- 1514263
- Publication, EPODOC
- EP1514263
- Application
- 3760242
- Application, DOCDB
- 03760242
- Application, EPODOC
- EP20030760242
Titles3
- German
- AUDIOCODIERUNGSSYSTEM, DAS EIGENSCHAFTEN EINES DECODIERTEN SIGNALS ZUR ANPASSUNG SYNTHETISIERTER SPEKTRALKOMPONENTEN VERWENDET
- English
- AUDIO CODING SYSTEM USING CHARACTERISTICS OF A DECODED SIGNAL TO ADAPT SYNTHESIZED SPECTRAL COMPONENTS
- French
- SYSTEME DE CODAGE AUDIO UTILISANT DES CARACTERISTIQUES D'UN SIGNAL DECODE POUR ADAPTER DES COMPOSANTS SPECTRAUX SYNTHETISES
Classification
- CPC, 1
- G10L21/038
- IPC, 4
- G10L21 02
- G10L19 02
- G10L19 00
- H03M7 30
Designated states1
- Contracting states, 1
- Türkiye
