Nova Patents
EP0446037A2

Hybrid perceptual audio coding.

Abstract

A hybrid coding technique for high quality coding of audio signals, using a subband filtering technique further refined to achieve a large number of subbands. Noise masking thresholds for subbands are then determined using a new tonality measure applicable to individual frequency bands or single frequencies. Based on the thresholds so determined, input signals are coded to achieve high quality at reduced bit rates.

EP0446037A2, drawing sheet 1
Sheet 1 of 16

Term

Term ended

Projected expiry passed 6 March 2011, 15.6 years ago.

  1. Priority
  2. Filed
  3. Published
  4. Projected expiry
  5. Today

22 claims: 4 independent, 18 dependent

  1. 1
    A method of processing an ordered time sequence of audio signals partitioned into blocks of samples, said method comprising    determining a discrete short-time spectrum, S(ω i ), i=1, 2,...,N, for each of said blocks,    determining the value of a tonality function as a function of frequency, and    based on said tonality function, estimating the noise masking threshold for each of ω i ,     CHARACTERIZED IN THAT    said step of determining S(ω i ) comprises determining S(ω i ) with differing time and frequency resolution as a function of ω i ,
  2. 11
    The method of any of claims 1, 8, 9 or 10, further     CHARACTERIZED IN THAT    said method further comprising the steps of    generating an estimate of the number of bits necessary to encode S(ω i )    quantizing said S(ω i ) to form quantized representations of said S(ω i ) using said estimate of the number of bits, and    providing to a medium a coded representation of said quantized values and information describing about how said quantized values were derived.
  3. 12
    A method for decoding an ordered sequence of coded signals comprising first code signals representing values of the frequency components corresponding to a block of values of an audio signal and second code signals representing information about how said first signals were derived to represent said audio signal with reduced perceptual error, said method comprising    using said second signals to determine quantizing levels for said audio signal which reflect a reduced level of perceptual distortion,    reconstructing quantized values for said frequency content of said audio signal in accordance with said quantizing levels, and    transforming said reconstructed quantized spectrum to recover an estimate of the audio signal,     CHARACTERIZED IN THAT    said frequency components have variable rime and frequency resolution.
  4. 18
    Apparatus for processing an ordered time sequence of audio signals partitioned into blocks of samples comprising,    means for determining a discrete short-time spectrum, S(ω i ), i=1, 2,...,N, for each of said blocks,    means for determining the value of a tonality function as a function of frequency, and    means for estimating the noise masking threshold for each of ω i , in response to said tonality function,    said means for determining further comprises    means for determining S(ω i ) with dithering time and frequency resolution at different values of ω i ,