IL180712A

Combining audio signals using auditory scene analysis

Abstract

A process for combining audio channels combines the audio channels to produce a combined audio channel and dynamically applies one or more of time, phase, and amplitude or power adjustments to the channels, to the combined channel, or to both the channels and the combined channel. One or more of the adjustments are controlled at least in part by a measure of auditory events in one or more of the channels and/or the combined channel. Applications include the presentation of multichannel audio in cinemas and vehicles. Not only methods, but also corresponding computer program implementations and apparatus implementations are included.

IL180712A, drawing sheet 1
Sheet 1 of 5

Term

No projected expiry on record.

  1. Priority
  2. Filed
  3. Published
  4. Today

18 claims: 4 independent, 14 dependent

  1. 1
    \ CLAIMS 1. A process for combining audio channels, comprising combining the audio channels to produce a combined audio channel, and dynamically applying one or more of time, phase, and amplitude or power adjustments to the channels, to the combined channel, or to both the channels and the combined channel, wherein one or more of said adjustments are controlled at least in part by a measure of auditory events in one or more of the channels and/or the combined channel so that the adjustments remain substantially constant during auditory events and are allowed to change at or near auditory event boundaries, wherein each auditory event boundary is identified in response to a change in signal characteristics with respect to time in a channel exceeding a threshold such that a set of auditory event boundaries is obtained for the channel, wherein an audio segment in the channel between consecutive boundaries constitutes an auditory event.
  2. 8
    A process for downmixing three input audio channels a, β, and δ to two output audio channels a and δ, wherein the three input audio channels represent, in order, consecutive spatial directions a, β, and δ, and the two output channels a and δ represent the non-consecutive spatial directions a and δ, comprising extracting common signal components from the two input audio channels representing directions a and δ to produce three intermediate channels:channel a׳, a modification of channel a representing the direction a, channel a׳ comprising the signal components of channel a from which signal components common to input channels a and δ have been substantially removed, channel δ׳, a modification of channel δ representing the direction δ, channel δ׳ comprising the signal components of channel δ from which signal components common to input channels a and δ have been substantially removed, and channel β׳, a new channel representing the direction β, channel β' comprising the signal components common to input channels a and δ, combining intermediate channel a׳, intermediate channel β׳, and input channel β to produce output channel a, and combining intermediate channel δ', intermediate channel β', and input channel β to produce output channel δ״.
  3. 12
    Apparatus adapted to perform the methods of any one of claims 1, 8 and 10.
  4. 13
    A computer program, stored on a computer-readable medium for causing a computer to perform the methods of any one of claims 1, 8 and 10.