IL102146A

Variable rate vocoder - a method and apparatus for signal compression

Abstract

An apparatus for masking frame errors comprising memory means for storing at least one previous frame of data and for providing said at least one previous frame of data in response to a frame error signal; and masking means for receiving said frame error signal and for generating a masking signal in accordance with at least one previous frame of data and a predetermined error masking format.

IL102146A, drawing sheet 1
Sheet 1 of 27

Term

No projected expiry on record.

  1. Priority
  2. Filed
  3. Published
  4. Today

28 claims: 2 independent, 26 dependent

  1. 1
    A method of speech signal compression, by variable rate coding of frames of digitized speech samples, comprising the steps of:determining a level of speech activity for a frame of digitized speech samples;selecting an encoding rate from a set of rates based upon said determined level of speech activity for said frame;coding said frame according to a coding format of a set of coding formats for said selected rate wherein each rate has a corresponding different coding format and wherein each coding format provides for a different plurality of parameter signals representing said digitized speech samples in accordance with a speech model;and generating for said frame a data packet of said parameter signals at said selected rate.
  2. 2
    The method of Claim 1 wherein said step of determining said level of frame speech activity comprises the steps of:determining said level of speech activity from said frame of digitized speech samples;comparing said frame speech activity with said at least one speech activity threshold level of a predetermined set of activity threshold levels;generating an indication when said frame speech activity exceeds each corresponding one of said at least one speech activity threshold levels;and adaptively adjusting at least one of said at least one speech activity threshold levels with respect to a level of activity of a previous frame of digitized speech samples.
  3. 3
    The method of Claim 1 further comprising the steps of:generating a rate command indicative of a preselected encoding rate for said frame;and modifying said selected encoding rate to provide said preselected encoding rate for coding of said frame at said preselected encoding rate.
  4. 4
    The method of Claim 3 wherein said preselected rate is less than a predetermined maximum rate, said method further comprising the steps of:- 102146 (5 generating an additional data packet;and combining said data packet with said additional data packet within a transmission frame for transmission.
  5. 5
    The method of Claim 1 wherein said step of generating said data packet of said parameter signals comprises:generating a variable number of bits to represent linear predictive coefficient (LPC) vector signals of said frame of digitized speech samples, wherein said variable number of bits representing said LPC vector signals is responsive to said measured speech activity level;generating a variable number of bits to represent pitch vector signals of said frame of digitized speech samples, wherein said variable number of bits representing said pitch vector signals is responsive to said measured speech activity level;and generating a variable number of bits to represent codebook excitation vector signals of said frame of digitized speech samples, wherein said variable number of bits representing said codebook excitation vector signals is responsive to said measured speech activity level.
  6. 6
    The method of Claim 1 wherein said step coding said frame comprises:generating for said frame a variable number of linear prediction coefficients wherein said variable number of said linear prediction coefficients is responsive to said selected encoding rate;generating for said frame a variable number of pitch coefficients wherein said variable number of said pitch coefficients is responsive to said selected encoding rate;and generating for said frame a variable number of codebook excitation values wherein said variable number of said codebook excitation values is responsive to said selected encoding rate.
  7. 7
    The method of Claim 1 wherein said step of determining a level of speech activity comprises summing the squares of the values of said digitized speech samples. 102146/^
  8. 8
    The method of Claim 7 further comprising the step of generating 2 error protection bits for said data packet.
  9. 9
    The method of Claim 8 wherein said step of generating error 2 protection bits for said data packet wherein the number of said protection bits is responsive to said frame of speech activity level.
  10. 10
    The method of Claim 1 wherein said step of adaptively adjusting 2 speech activity threshold levels comprises the steps of:comparing said measured speech activity to said at least one of speech 4 activity thresholds and incrementally increasing said at least one of speech activity thresholds toward the level of said frame speech activity when said 6 frame speech activity exceeds said at least one of said speech activity thresholds;and 8 comparing said measured speech activity to said at least one of speech activity thresholds and decreasing said at least one of speech activity 10 thresholds to the level of said frame speech activity when said frame speech activity is less than said at least one of speech activity thresholds.
  11. 11
    The method of Claim 8 wherein said step of generating error 2 protection for said data packet further comprises determining the values of said error protection bits in accordance with a cyclic block code.
  12. 12
    The method of Claim 10 wherein said step of selecting an 2 encoding rate is responsive to an external rate signal.
  13. 13
    The method of Claim 1 further comprising the step of pre2 multiplying said digitized speech samples by a predetermined windowing function.
  14. 14
    The method of Claim 1 further comprising the step of 2 converting said LPC coefficients to line spectral pair (LSP) values. ^102146/2 ־74־
  15. 15
    The method of Claim 1 wherein said input frame of digitized samples comprises digitized values for approximately twenty milliseconds of speech.
  16. 16
    The method of Claim 1 wherein said input frame of digitized samples comprises approximately 160 digitized samples.
  17. 17
    The method of Claim 1 wherein said output data packet comprises:one hundred and seventy one bits comprised of forty bits for LPC data, forty bits for pitch data, eighty bits for excitation vector data and eleven bits for error protection when said output data rate is full rate;eighty bits comprised of twenty bits for LPC information, twenty bits for pitch information and forty bits for excitation vector data when said output data rate is half rate;forty bits comprised of ten bits for LPC information, ten bits for pitch information and twenty bits for excitation vector data when said output data rate is quarter rate;and sixteen bits comprised of ten bits for LPC information and six bits for excitation vector information when said output data rate is eighth rate.
  18. 18
    An apparatus for compressing an acoustical signal into variable rate data comprising:means for determining a level of audio activity for an input frame of digitized samples of said acoustical signal;means for selecting an output data rate from a predetermined set of rates based upon said determined level of audio activity within said frame;means for coding said frame according to a coding format of a set of coding formats for said selected rate wherein each rate has a corresponding different coding format to provide a plurality of parameter signals representing said digitized speech samples in accordance with a speech model;and means for providing for said frame a corresponding data packet at a data rate corresponding to said selected rate wherein each coding format of data packets at different rates represent a different plurality of parameter signals representing said digitized speech samples in accordance with said speech model.
  19. 19
    The apparatus for compressing an acoustical signal of Claim 18 wherein said output data packet comprises:a variable number of bits to represent LPC vector signals of said frame of digitized speech samples, wherein said variable number of bits for representing said LPC vector signals is responsive to said level of audio activity;a variable number of bits to represent pitch vector signals of said frame of digitized speech samples, wherein said variable number of bits for representing said pitch vector signals is responsive to said level of audio activity;and a variable number of bits to represent codebook excitation vector signals of said frame of digitized speech samples, wherein said variable number of bits for representing said codebook excitation vector signals is responsive to said level of audio activity.
  20. 20
    The apparatus for compressing an acoustical signal of Claim 18 wherein said means for determining said level of audio activity comprises:means for determining an energy value for said input frame;means for comparing said input frame energy with said at least one audio activity thresholds;and means for generating an indication when said input frame activity exceeds each corresponding one of said at least one audio activity thresholds.
  21. 21
    The apparatus of Claim 18 wherein said means for determining said energy of said input frame comprises:squaring means for squaring said digitized audio samples of a frame;and summing means for summing said squares of digitized audio samples of a frame.
  22. 22
    The apparatus of Claim 18 wherein said means for determining a level of audio activity comprises:- 1021*46/3 means for calculi!ling a set of linear predictive coefficients for said 4 input frame of digitized samples of said acoustical signals;and means for determining said level of audio activity in accordance with 6 al least one of said linear predictive coefficients.
  23. 23
    The apparatus of Claim 18 further comprising means for 2 providing error protection bits for said data packet responsive to said.selected ־ output data rate.
  24. 24
    The apparatus of Claim 18 further comprising a means for 2 converting said LPC coefficients to line spectral pair (LSP) values.
  25. 25
    The apparatus of Claim 19 further comprising a means for 2 adaptively adjusting said at least one of said at least one audio activity thresholds.
  26. 26
    The apparatus of Claim 23 wherein said means for providing 2 error protection bits provides the values of said error protection bits in accordance with a cyclic block code.
  27. 27
    The apparatus of Claim 19 wherein said set of rales comprises 2 full rale, half rale, quarter rale and eighth rale.
  28. 28
    I he apparatus of Claim 19 wherein said set of rales comprises 2 8 Kbps, 4 Kbps, 2 Kbps and !Kbps. I g J
Independent claims28