CA2964368C

Jitter buffer control, audio decoder, method and computer program

Abstract

A jitter buffer control for controlling a provision of a decoded audio content on the basis of an input audio content is configured to select a frame-based time scaling or a sample-based time scaling in a signal-adaptive manner. An audio decoder uses such a jitter buffer control.

CA2964368C, drawing sheet 1
Sheet 1 of 14

Term

7.7 yearsleft in the term

Expires 18 June 2034.

  1. Priority
  2. Filed
  3. Granted
  4. Today
  5. Expires

3 claims: 2 independent, 1 dependent

  1. 1
    Claims 1. A jitter buffer control for controlling a provision of a decoded audio content on the basis of an input audio content. wherein the jitter buffer control is configured to select a frame-based time scaling or a sample-based time scaling in a signal-adaptive manner;wherein audio frames are dropped or inserted to control a depth of a jitter buffer when the frame-based time scaling is used, and wherein a time-shifted overlapand-add of audio signal portions is performed when the sample-based time-scaling is used;wherein the jitter buffer control is configured to select a frame-based comfort noise insertion or a frame-based comfort noise deletion for a time scaling if a discontinuous transmission in conjunction with comfort noise generation is currently used or was used for a previous frame, wherein the jitter buffer control is configured to select an overlap-add-operation using a predetermined time shift for the time scaling if a current audio signal portion is active but comprises a signal energy which is smaller than or equal to an energy threshold value, and if a jitter buffer is not empty, or if a previous audio signal portion was active but comprises a signal energy which is smaller than or equal to the energy threshold value, and if the jitter buffer is not empty;wherein the jitter buffer control is configured to select an overlap-add-operation using a signal-adaptive time shift for the time scaling if the current audio signal portion is active and comprises a signal energy which is larger than or equal to the energy threshold value and if the jitter buffer is not empty, or if the previous audio signal portion was active and comprises a signal energy which is larger than or equal to the energy threshold value and if the jitter buffer is not empty;and wherein the jitter buffer control is configured to select an insertion of a concealed frame for the time scaling if the current audio signal portion is active and if the jitter buffer is empty, or if the previous audio signal portion was active and if the jitter buffer is empty. CA 2964368 2019-04-02
  2. 2
    A method for controlling a provision of a decoded audio content on the basis of an input audio content, wherein the method comprises selecting a frame-based time scaling or a samplebased time scaling in a signal-adaptive manner;wherein audio frames are dropped or inserted to control a depth of a jitter buffer when the frame-based time scaling is used, and wherein a time-shifted overlapand-add of audio signal portions is performed when the sample-based time-scaling is used;wherein the method comprises selecting a frame-based comfort noise insertion or a frame-based comfort noise deletion for a time scaling if a discontinuous transmission in conjunction with comfort noise generation is currently used or was used for a previous frame, selecting an overlap-add-operation using a predetermined time shift for the time scaling if a current audio signal portion is active but comprises a signal energy which is smaller than or equal to an energy threshold value, and if a jitter buffer is not empty, or if a previous audio signal portion was active but comprises a signal energy which is smaller than or equal to the energy threshold value, and if the jitter buffer is not empty;selecting an overtap-add-operation using a signal-adaptive time shift for the time scaling if the current audio signal portion is active and comprises a signal energy which is larger than or equal to the energy threshold value and if the jitter buffer is not empty, or if the previous audio signal portion was active and comprises a signal energy which is larger than or equal to the energy threshold value and if the jitter buffer is not empty;and selecting an insertion of a concealed frame few the time scaling if the current audio signal portion is active and if the jitter buffer is empty, or if the previous audio signal portion was active and if the jitter buffer is empty. CA 2964368 2019-04-02