CA2898677C

Low-frequency emphasis for lpc-based coding in frequency domain

Abstract

The invention provides an audio encoder and method for encoding a non-speech audio signal so as to produce therefrom a bitstream, the audio encoder comprising: a combination (2, 3) of a linear predictive coding filter (2) having a plurality of linear predictive coding coefficients (LC) and a time-frequency converter (3), wherein the combination (2, 3) is configured to filter and to convert a frame (Fl) of the audio signal (AS) into a frequency domain in order to output a spectrum (SP) based on the frame (Fl) and on the linear predictive coding coefficients (LC); a low frequency emphasizer (4) configured to calculate a processed spectrum (PS) based on the spectrum (SP), wherein spectral lines (SL) of the processed spectrum (PS) representing a lower frequency than a reference spectral line (RSL) are emphasized; and a control device (5) configured to control the calculation of the processed spectrum (PS) by the low frequency emphasizer (4) depending on the linear predictive coding coefficients (LC) of the linear predictive coding filter (2). Furthermore, the invention provides a corresponding audio decoder, a system, a method for decoding a bitstream containing quantized spectrums and a plurality of linear predictive coding coefficients and a corresponding computer program.

CA2898677C, drawing sheet 1
Sheet 1 of 11

Term

7.3 yearsleft in the term

Expires 28 January 2034.

  1. Priority
  2. Filed
  3. Granted
  4. Today
  5. Expires

39 claims: 4 independent, 35 dependent

  1. 1
    CA 02898677 2016-12-02 Claims 1. Audio encoder for encoding a non-speech audio signal so as to produce therefrom a bitstream, the audio encoder comprising:a combination of a linear predictive coding filter having a plurality of linear predictive coding coefficients and a time-frequency converter, wherein the combination is configured to filter and to convert a frame of the nonspeech audio signal into a frequency domain in order to output a spectrum based on the frame and on the linear predictive coding coefficients ;a low frequency emphasizer configured to calculate a processed spectrum based on the spectrum, wherein spectral lines of the processed spectrum representing a lower frequency than a reference spectral line are emphasized;a control device configured to control the calculation of the processed spectrum by the low frequency emphasizer depending on the linear predictive coding coefficients of the linear predictive coding filter: a quantization device configured to produce a quantized spectrum based on the processed spectrum;and a bitstream producer configured to embed the quantized spectrum and the linear predictive coding coefficients into the bitstream.
  2. 4
    Audio encoder according to any one of claims 1 to 3, wherein the control device comprises a spectral analyzer configured to estimate a spectral representation of the linear predictive coding coefficients, a minimummaximum analyzer configured to estimate a minimum of the spectral representation and a maximum of the spectral representation below a further reference spectral line and an emphasis factor calculator configured to calculate spectral line emphasis factors for calculating the spectral lines of the processed spectrum representing a lower frequency than the reference spectral line based on the minimum and on the maximum, wherein the spectral lines ofthe processed spectrum are emphasized by applying the spectral line emphasis factors to spectral lines of the spectrum of the filtered frame.
  3. 6
    Audio encoder according to any one of claims 4 or 5, wherein the emphasis factor calculator comprises a first stage configured to calculate a basis emphasis factor according to a first formula y = (a · min / max) 13 , wherein a is a first preset value, with a > 1, β is a second preset value, with 0 < β < 1, min is the minimum ofthe spectral representation, max is the maximum of the spectral representation and γ is the basis emphasis factor, and wherein the emphasis factor calculator comprises a second stage CA 02898677 2016-12-02 configured to calculate spectral line emphasis factors according to a second formula ει = γ Ν , wherein i’ is a number of the spectral lines to be emphasized, i is an index of the respective spectral line, the index increases with the frequencies of the spectral lines, with i = 0 to i’-1, γ is the basis emphasis factor and ει is the spectral line emphasis factor with index i.
  4. 10
    Audio encoder according to any one of claims 6 to 9, wherein the second preset value is determined according to the formula β = 1 / (Θ · i j, wherein i’ is the number of the spectral lines being emphasized, Θ is a factor between 3 and 5.
  5. 11
    Audio encoder according to any one of claims 6 to 9, wherein the second preset value is determined according to the formula β = 1 / (θ · i j, wherein i’ is the number of the spectral lines being emphasized, θ is a factor between 3,4 and 4,6.
  6. 12
    Audio encoder according to any one of claims 6 to 9, wherein the second preset value is determined according to the formula β = 1 / (Θ i j, wherein i’ is the number of the spectral lines being emphasized, θ is a factor between 3,8 and 4,2.
  7. 13
    Audio encoder according to any one of claims 1 to 12, wherein the reference spectral line represents a frequency between 600 Hz and 1000Hz. CA 02898677 2016-12-02
  8. 14
    Audio encoder according to any one of claims 1 to 12, wherein the reference spectral line represents a frequency between 700 Hz and 900 Hz.
  9. 15
    Audio encoder according to any one of claims 1 to 12, wherein the reference spectral line represents a frequency between 750 Hz and 850 Hz.
  10. 16
    Audio encoder according to any one of claims 4 to 15, wherein the further reference spectral line represents the same or a higher frequency than the reference spectral line.
  11. 17
    Audio encoder according to any one of claims 1 to 16, wherein the control device is configured in such way that the spectral lines of the processed spectrum representing a lower frequency than the reference spectral line are emphasized only if the maximum is less than the minimum multiplied with the first preset value.
  12. 18
    Audio decoder for decoding a bitstream based on a non-speech audio signal so as to produce from the bitstream a non-speech audio output signal, the bitstream containing quantized spectrums and a plurality of linear predictive coding coefficients, the audio decoder comprising:a bitstream receiver configured to extract the quantized spectrum and the linear predictive coding coefficients from the bitstream;a de-quantization device configured to produce a de-quantized spectrum based on the quantized spectrum;a low frequency de-emphasizer configured to calculate a reverse processed spectrum based on the de-quantized spectrum, wherein spectral lines of the reverse processed spectrum representing a lower frequency than a reference spectral line are deemphasized;and CA 02898677 2016-12-02 a control device configured to control the calculation of the reverse processed spectrum by the low frequency de-emphasizer depending on the linear predictive coding coefficients contained in the bitstream.
  13. 22
    Audio decoder according to any one of claims 18 to 21, wherein the control device comprises a spectral analyzer configured to estimate a spectral representation of the linear predictive coding coefficients, a minimummaximum analyzer configured to estimate a minimum of the spectral representation and a maximum of the spectral representation below a further reference spectral line and a de-emphasis factor calculator configured to CA 02898677 2016-12-02 calculate spectral line de-emphasis factors for calculating the spectral lines of the reverse processed spectrum representing a lower frequency than the reference spectral line based on the minimum and on the maximum, wherein the spectral lines of the reverse processed spectrum are de-emphasized by applying the spectral line de-emphasis factors to spectral lines of the spectrum of the de-quantized spectrum.
  14. 24
    Audio decoder according to any one of claims 22 or 23, wherein the deemphasis factor calculator comprises a first stage configured to calculate a basis de-emphasis factor according to a first formula δ = (a min / max)' P, wherein a is a first preset value, with a > 1, β is a second preset value, with 0 < β < 1, min is the minimum of the spectral representation, max is the maximum of the spectral representation and δ is the basis de-emphasis factor, and wherein the de-emphasis factor calculator comprises a second stage configured to calculate spectral line de-emphasis factors according to a second formula ζί = δ' -', wherein ί’ is a number of the spectral lines to be de-emphasized, i is an index ofthe respective spectral line, the index increases with the frequencies of the spectral lines, with i = 0 to i’-1, δ is the basis de-emphasis factor and ζί is the spectral line de-emphasis factor with index i.
  15. 28
    Audio decoder according to any one of claims 24 to 27, wherein the second preset value is determined according to the formula β = 1 / (Θ i’), wherein i’ is the number of the spectral lines being de-emphasized, Θ is a factor between 3 and 5.
  16. 29
    Audio decoder according to any one of claims 24 to 27, wherein the second preset value is determined according to the formula β = 1 / (θ i’), wherein i’ is the number ofthe spectral lines being de-emphasized, θ is a factor between 3,4 and 4,6.
  17. 30
    Audio decoder according to any one of claims 24 to 27, wherein the second preset value is determined according to the formula β = 1 / (Θ i’), wherein i’ is the number ofthe spectral lines being de-emphasized, θ is a factor between 3,8 and 4,2.
  18. 31
    Audio decoder according to any one of claims 18 to 30, wherein the reference spectral line represents a frequency between 600 Hz and 1000Hz.
  19. 32
    Audio decoder according to any one of claims 18 to 30, wherein the reference spectral line represents a frequency between 700 Hz and 900 Hz.
  20. 33
    Audio decoder according to any one of claims 18 to 30, wherein the reference spectral line represents a frequency between 750 Hz and 850 Hz.
  21. 34
    Audio decoder according to any one of claims 22 to 33, wherein the further reference spectral line represents the same or a higher frequency than the reference spectral line. CA 02898677 2016-12-02
  22. 35
    Audio decoder according to any one of claims 18 to 34, wherein the control device is configured in such way that the spectral lines of the reverse processed spectrum representing a lower frequency than the reference spectral line are de-emphasized only if the maximum is less than the minimum multiplied with the first preset value.
  23. 36
    A system comprising a decoder and an encoder, wherein the encoder is designed according to any one of claims 1 to 17 and the decoder is designed according to any one of claims 18 to 35.
  24. 37
    Method for encoding a non-speech audio signal so as to produce therefrom a bitstream, the method comprising the steps:filtering with a linear predictive coding filter having a plurality of linear predictive coding coefficients and converting a frame of the non-speech audio signal into a frequency domain in order to output a spectrum based on the frame and on the linear predictive coding coefficients;calculating a processed spectrum based on the spectrum, wherein spectral lines of the processed spectrum representing a lower frequency than a reference spectral line are emphasized;and controlling the calculation of the processed spectrum depending on the linear predictive coding coefficients of the linear predictive coding filter;producing a quantized spectrum based on the processed spectrum;and embedding the quantized spectrum and the linear predictive coding coefficients into the bitstream.
  25. 38
    Method for decoding a bitstream based on a non-speech audio signal so as to produce from the bitstream a non-speech audio output signal, the CA 02898677 2016-12-02 bitstream containing quantized spectrums and a plurality of linear predictive coding coefficients, the method comprising the steps:extracting the quantized spectrum and the linear predictive coding coeffi5 cients from the bitstream;producing a de-quantized spectrum based on the quantized spectrum;calculating a reverse processed spectrum based on the de-quantized io spectrum, wherein spectral lines of the reverse processed spectrum representing a lower frequency than a reference spectral line are deemphasized;and controlling the calculation of the reverse processed spectrum depending 15 on the linear predictive coding coefficients contained in the bitstream.
  26. 39
    A computer-readable medium having computer-readable code stored thereon that when executed by a computer perform the method according to any one of claims 37 or 38.
Independent claims26