IL210331A

Low bitrate audio encoding/decoding scheme having cascaded switches

Abstract

This record has no abstract on file.

IL210331A, drawing sheet 1
Sheet 1 of 22

Term

No projected expiry on record.

  1. Priority
  2. Filed
  3. Published
  4. Today

25 claims: 21 independent, 4 dependent

  1. 1
    Audio encoder for encoding an audio input signal (195), the audio input signal being in a first domain, comprising:a first coding branch (400) for encoding an audio signal using a first coding algorithm to obtain a first encoded signal;a second coding branch (500) for encoding an audio signal using a second coding algorithm to obtain a second encoded signal, wherein the first coding algorithm is different from the second coding algorithm;and a first switch (200) for switching between the first coding branch and the second coding branch so that, for a portion of the audio input signal, either the first encoded signal or the second encoded signal is in an encoder output signal, wherein the second coding branch comprises: a converter (510) for converting the audio signal into a second domain different from the first domain, a first processing branch (522) for processing an audio signal in the second domain to obtain a first processed signal;a second processing branch (523, 524) for converting a signal into a third domain different from the first domain and the second domain and for processing the signal in the third domain to obtain a second processed signal;and a second switch (521) for switching between the first processing branch (522) and the second processing branch ¢523, 524) so that, for a portion of the audio signal input into the second 5 coding branch, either the first processed signal or the second processed signal is in the second encoded signal.
  2. 2
    Audio encoder in accordance with claim 1, in which the 10 first coding algorithm in the first coding branch (400) is based on an information sink model, or in which the second coding algorithm in the second coding branch (500) is based on an information source or a signal to noise ratio (SNR) model.
  3. 3
    Audio encoder in accordance with claim 1 or 2, in which the first coding branch comprises a converter (410) for converting the audio input signal into a fourth domain different from the first domain, the 20 second domain, and the third domain.
  4. 4
    Audio encoder in accordance with one of the preceding claims, in which the first domain is the time domain, the second domain is an LPC domain obtained by an LPC 25 filtering the first domain signal, the third domain is an LPC spectral domain obtained by converting an LPC filtered signal into a spectral domain, and the fourth domain is a spectral domain obtained by frequency domain converting the first domain signal.
  5. 5
    Audio encoder in accordance with one of the preceding claims, further comprising a controller (300, 525) for controlling the first switch (200) or the second switch (521) in a signal adaptive way, wherein the controller is operative to analyze a signal input into the first switch (200) or output by the first coding branch or the second coding branch or a signal obtained by decoding an output signal of the first coding branch or the second coding branch with respect to a target function, or wherein the controller (300, 525) is operative to analyze a signal input into the second switch (521) or output by the first processing branch or the second processing branch or signals obtained by inverse processing output signals from the first processing branch (522) and the second processing branch (523, 524) with respect to a target function.
  6. 6
    Audio encoder in accordance with one of the preceding claims, in which the first coding branch (400) or the second processing branch (523, 524) of the second coding branch (500) comprises an aliasing introducing time/frequency converter and a quantizer/entropy coder stage (421) and wherein the first processing branch of the second coding branch comprises a quantizer or entropy coder stage (522) without an aliasing introducing conversion.
  7. 7
    Audio encoder in accordance with claim 6, in which the aliasing introducing time/frequency converter comprises a windower for applying an analysis window and a modified discrete cosine transform (MDCT) algorithm, the windower being operative to apply the window function to subsequent frames in an overlapping manner so that a sample of an input signal into the windower occurs in at least two subsequent frames.
  8. 8
    Audio encoder in accordance with one of the preceding claims, in which the first processing branch (522) comprises the LPC excitation coding of an algebraic code excited linear prediction (ACELP) coder and the second processing branch comprises an MDCT spectral converter and a quantizer for quantizing spectral components to obtain quantized spectral components, wherein each quantized spectral component is zero or is defined by one quantization index of a plurality of quantization indices.
  9. 9
    Audio encoder in accordance with claim 5, in which the controller is operative to control the first switch (200) in an open loop manner and to control the second switch (521) in a closed loop manner.
  10. 10
    Audio encoder in accordance with one of the preceding claims, in which the first coding branch and the second coding branch are operative to encode the audio signal in a block wise manner, wherein the first switch or the second switch are switching in a blockwise manner so that a switching action takes place, at the minimum, after a block of a predefined number of samples of a signal, the predefined number of samples forming a frame length for the corresponding switch (521, 200).
  11. 11
    Audio encoder in accordance with claim 10, in which the frame length for the first switch is at least double the size of the frame length of the second switch.
  12. 12
    Audio encoder in accordance with claim 5, in which the controller is operative to perform a speech/music discrimination in such a way that a decision to speech is favored with respect to a decision to music so that a decision to speech is taken even when a portion less than 50% of a frame for the first switch is speech and a portion more than 50% of the frame for the first switch is music.
  13. 13
    Audio encoder in accordance with claim 5 or 12, in which a frame for the second switch is smaller than a frame for the first switch, and in which the controller (525, 300) is operative to take a decision to speech when only a portion of the first frame which has a length which is more than 50% of the length of the second frame is found out to include music.
  14. 14
    Audio encoder in accordance with one of the preceding claims, in which the first encoding branch (400) or the second processing branch of the second coding branch includes a variable time warping functionality.
  15. 15
    Method of encoding an audio input signal (195), the audio input signal being in a first domain, comprising:encoding (400) an audio signal using a first coding algorithm to obtain a first encoded signal;encoding (500) an audio signal using a second coding algorithm to obtain a second encoded signal, wherein the first coding algorithm is different from the second coding algorithm;and switching (200) between encoding using the first coding algorithm and encoding using the second coding algorithm so that, for a portion of the audio input signal, either the first encoded signal or the second encoded signal is in an encoded output signal, wherein encoding (500) using the second coding algorithm comprises : converting (510) the audio signal into a second domain different from the first domain, processing (522) an audio signal in the second domain to obtain a first processed signal;converting (523) a signal into a third domain different from the first domain and the second domain and processing (524) the signal in the third domain to obtain a second processed signal;and switching (521) between processing (522) the audio signal and converting (523) and processing (524) so that, for a portion of the audio signal encoded using the second coding algorithm, either the first processed signal or the second processed signal is in the second encoded signal.
  16. 16
    Decoder for decoding an encoded audio signal, the encoded audio signal comprising a first coded signal, a first processed signal in a second domain, and a second processed signal in a third domain, wherein the first coded signal, the first processed signal, and the second processed signal are related to different time portions of a decoded audio signal, and wherein a first domain, the second domain and the third domain are different from each other, comprising:a first decoding branch (431, 440) for decoding the first encoded signal based on the first coding algorithm;a second decoding branch for decoding the first processed signal or the second processed signal, wherein the second decoding branch comprises a first inverse processing branch (531) for inverse processing the first processed signal to obtain a first inverse processed signal in the second domain;a second inverse processing branch (533, 534) for inverse processing the second processed signal to obtain a second inverse processed signal in the second domain;a first combiner (532) for combining the first inverse processed signal and the second inverse processed signal to obtain a combined signal in the second domain;and a converter (540) for converting the combined signal to the first domain;and a second combiner (600) for combining the converted signal in the first domain and the first decoded signal output by the first decoding branch to obtain a decoded output signal in the first domain.
  17. 17
    Decoder of the claim 16, in which the first combiner (532) or the second combiner (600) comprises a switch having a cross fading functionality.
  18. 22
    Decoder in accordance with one of claims 16 to 21, in which the encoded signal comprises, as side information (4a), an indication whether a coded signal is to be coded by a first encoding branch or a second encoding branch or a first processing branch of the second encoding branch or a second processing branch of the second encoding branch, and which further comprises a parser for parsing the encoded signal to determine, based on the side information (4a), whether a coded signal is to be processed by the first decoding branch, or the second decoding branch, or the first inverse processing branch of the second decoding branch or the second inverse processing branch of the second decoding branch.
  19. 23
    Method of decoding an encoded audio signal, the encoded audio signal comprising a first coded signal, a first processed signal in a second domain, and a second processed signal in a third domain, wherein the first coded signal, the first processed signal, and the second processed signal are related to different time portions of a decoded audio signal, and wherein a first domain, the second domain and the third domain are different from each other, comprising:decoding (431, 440) the first encoded signal based on a first coding algorithm;decoding the first processed signal or the second processed signal, wherein the decoding the first processed signal or the second processed signal comprises: inverse processing (531) the first processed signal to obtain a first inverse processed signal in the second domain;inverse processing (533, 534) the second processed signal to obtain a second inverse processed signal in the second domain;combining ¢532) the first inverse processed signal and the second inverse processed signal to obtain a combined signal in the second domain;and converting (540) the combined signal to the first domain;and combining (600) the converted signal in the first domain and the decoded first signal to obtain a decoded output signal in the first domain.
  20. 24
    Encoded audio signal comprising:a first coded signal encoded or to be decoded using a first coding algorithm, a first processed signal in a second domain, and a second processed signal in a third domain, wherein the first processed signal and the second processed signal are encoded using a second coding algorithm, wherein the first coded signal, the first processed signal, and the second processed signal are related to different time portions of a decoded audio signal, wherein a first domain, the second domain and the third domain are different from each other, and side information (4a) indicating whether a portion of the encoded signal is the first coded signal, the first processed signal or the second processed signal.
  21. 25
    Computer program for performing, when running on the computer, the method of encoding an audio signal in accordance with claim 15 or the method of decoding an encoded audio signal in accordance with claim 23.
Independent claims21