Nova Patents
MY199032A

Audio encoder and decoder

Abstract

The present disclosure provides methods, devices and computer program products for encoding and decoding of a vector (114, 902, 1002) of parameters in an audio coding system. The disclosure further relates to a method and apparatus for reconstructing an audio object in an audio decoding system. According to the disclosure, a modulo differential approach for coding and encoding a vector of a non-periodic quantity may improve the coding efficiency and provide encoders and decoders with less memory requirements. Moreover, an efficient method for encoding and decoding a sparse matrix is provided.

MY199032A, drawing sheet 1
Sheet 1 of 8

Term

No projected expiry on record.

  1. Priority and filed
  2. Published
  3. Today

18 claims: 4 independent, 14 dependent

  1. 1
    CLAIMS 1. A method for encoding an upmix matrix in an audio encoding system, each row of the upmix matrix comprising M elements allowing reconstruction of a time/frequency tile of an audio object from a downmix signal comprising M channels, the method comprising:for each row in the upmix matrix: selecting a subset of elements from the M elements of the row in the upmix matrix, wherein the selected subset of elements comprises a same number of elements for each row of the upmix matrix;representing each element in the selected subset of elements by a value and a position in the upmix matrix;and encoding the value and the position in the upmix matrix of each element in the selected subset of elements.
  2. 9
    An encoder (100) for encoding an upmix matrix in an audio encoding system, each row of the upmix matrix comprising M elements allowing reconstruction of a time/frequency tile of an audio object from a downmix signal comprising M channels, the encoder comprising:a receiving component adapted to receive each row in the upmix matrix;a selection component adapted to select a subset of elements from the M elements of the row in the upmix matrix, wherein the selected subset of elements comprises a same number of elements for each row of the upmix matrix;and an encoding component adapted to represent each element in the selected subset of elements by a value and a position in the upmix matrix, the encoding further adapted to encode the value and the position in the upmix matrix of each element in the selected subset of elements.
  3. 10
    A method for reconstructing a plurality of time/frequency tiles of an audio object in an audio decoding system (1200), comprising, for each time/frequency tile:receiving a downmix signal (1210) comprising M channels;receiving at least one encoded element (1204) representing a subset of M elements of a row in an upmix matrix, each encoded element comprising a value and a position in the row in the upmix matrix, the position indicating one of the M channels of the downmix signal to which the encoded element corresponds;and reconstructing (1208) the time/frequency tile of the audio object from the downmix signal by forming a linear combination of the downmix channels that correspond to the at least one encoded element, wherein in said linear combination each downmix channel is multiplied by the value of its corresponding encoded element, - 32 wherein the at least one encoded element comprises a same number of elements for each time/frequency tile.
  4. 18
    A decoder (1200) for reconstructing a plurality of time/frequency tiles of an audio object, comprising, for each time/frequency tile:a receiving component (1206) configured to receive a downmix signal (1210) comprising M channels and at least one encoded element (1204) representing a subset of M elements of a row in an upmix matrix, each encoded element comprising a value and a position in the row in the upmix matrix, the position indicating one of the M channels of the downmix signal to which the encoded element corresponds;and a reconstructing component (1208) configured to reconstruct the time/frequency tile of the audio object from the downmix signal by forming a linear combination of the downmix channels that correspond to the at least one encoded element, wherein in said linear combination each downmix channel is multiplied by the value of its corresponding encoded element, wherein the at least one encoded element comprises a same number of elements for each time/frequency tile.