US9620137B2

Determining between scalar and vector quantization in higher order ambisonic coefficients

Summary by NHIP

HOA Quantization Decoder

The method decodes bitstreams containing higher-order ambisonic coefficients by selecting between vector or scalar dequantization based on a syntax element. Vector dequantization determines N weight values from a codebook, where M greatest values are selected to reconstruct the spatial component.

Claim Score by NHIP

Read claim 19, the broadest

Abstract

In general, techniques are described for coding of vectors decomposed from higher-order ambisonic coefficients. A device comprising a memory and a processor may perform the techniques. The memory may be configured to store audio data. The processor may be configured to determine whether to perform vector dequantization or scalar dequantization with respect to a decomposed version of the plurality of HOA coefficients.

US9620137B2, drawing sheet 1
Sheet 1 of 64

Term

8.7 yearsleft in the term

Expires 29 May 2035, including 15 days of term adjustment.

  1. Priority
  2. Filed
  3. Granted
  4. Today
  5. Expires

20 claims: 3 independent, 17 dependent

  1. 1
    A method of decoding a bitstream indicative of a plurality of higher-order ambisonic (HOA) coefficients representative of a soundfield, the method comprising:obtaining, by an audio decoding device, the bitstream, wherein the bitstream includes a syntax element identifying whether the vector quantization or the scalar quantization was performed;performing, by the audio decoding device and based on the syntax element identifying whether the vector quantization or the scalar quantization was performed, either vector dequantization or scalar dequantization with respect to a spatial component defined in a spherical harmonic domain;reconstructing, by the audio decoding device, the plurality of HOA coefficients based on the dequantized spatial component;rendering, by the audio decoding device, one or more loudspeaker feeds based on the reconstructed plurality of HOA coefficients;andreproducing, by one or more loudspeakers coupled to the audio decoding device, the soundfield based on the one or more loudspeaker feeds.
  2. 10
    A device configured to decode a bitstream indicative of a plurality of higher-order ambisonic (HOA) coefficients representative of a soundfield, the device comprising:a memory configured to store the bitstream that includes a syntax element that identifies whether the vector quantization or the scalar quantization was performed;andone or more processors coupled to the memory, and configured to: perform, based on the syntax element that identifies whether the vector quantization or the scalar quantization was performed, either vector dequantization or scalar dequantization with respect to a spatial component defined in a spherical harmonic domain;reconstruct the plurality of HOA coefficients based on the dequantized spatial component;andrender one or more loudspeaker feeds based on the reconstructed plurality of HOA coefficients;andone or more loudspeakers coupled to the processor, and configured to reproduce the soundfield based on the one or more loudspeaker feeds.
  3. 19
    Broadest claimClaim Score 61, broad(NHIP)A method of encoding audio data indicative of a plurality of higher-order ambisonic (HOA) coefficients representative of a soundfield, the method comprising:capturing, by a microphone coupled to an audio encoding device, the audio data;anddetermining, by the audio encoding device, whether to perform vector quantization or scalar quantization with respect to a spatial component decomposed from the plurality of HOA coefficients;performing, by the audio encoding device and so as to generate a bitstream including an encoded version of the audio data, either the vector quantization or the scalar quantization with respect to the spatial component based on the determination;andspecifying, by the audio encoding device and in the bitstream, a syntax element indicating whether the vector quantization or the scalar quantization was performed.