Nova Patents
US3158685A

Synthesis of speech from code signals

Abstract

This record has no abstract on file.

US3158685A, drawing sheet 1
Sheet 1 of 14

Term

Term ended

Expired 24 November 1981, 44.8 years ago.

  1. Priority and filed
  2. Granted
  3. Expired
  4. Today

8 claims: 8 independent, 0 dependent

  1. 1
    What is claimed is:1. Apparatus for the production of artificial speech-like sounds which comprises, a source of speech phoneme representations ordered according to a desired phonetic sequence, means for selectively analyzing selected sequences of said phoneme representations to produce a code signal for each individual phoneme in said sequence and for the vowel-consonant structure of said phoneme sequence associated with each inidividual phoneme, means responsive to said code signals representative of successive phonemes and associated vowel-consonant structures for developing a set of speech defining signals which together specify the acoustic parameters of said sequence of phonemes and the transitions between phonemes as a function of time, and synthesizer means continuously supplied with all of said speech defining signals for generating artificial speech-like sounds.
  2. 2
    In a mechanism for producing speech-like sounds, the combination which comprises:means for storing a set of speech parameters for each of a number of speech sounds;means responsive to a selected sequence of said stored parameters for generating a first sequence of control signals, each of which is representative of one of said speech sounds;means responsive to successive pairs of said stored parameters for generating a second sequence of control signals which vary in a substantially linear fashion between pairs of control signals of said first sequence that represent consecutive vowel or consecutive consonant speech sounds;means responsive to successive pairs of said stored parameters for generating a third sequence of control signals which vary in a substantially nonlinear fashion between pairs of control signals of said first signal sequence that represent consecutive vowelconsonant, or consecutive consonant-vowel speech sounds;the control signals of said first, second, and third sequences thus together representing the acoustic parameters of said speech sounds and the transitions uniquely associated with successive pairs of said speech sounds;and means responsive to all of said control signals together for generating artificial speech.
  3. 3
    Apparatus for the production of artificial speech-like sounds which comprises:a source of coded representations of phonemes of speech ordered according to a de sired phonetic sequence, means responsive to a succession of said representations for generating control signals that persist with a constant, specified, magnitude for intervals in said succession of representations which denote discrete phonemes, means responsive to a succession of said representations for generating control signals that vary both in magnitude and duration according to a prescribed schedule for intervals in said succession which do not denote discrete phonemes but which are bounded by such phoneme intervals, and means for utilizing said control signal in the generation of artificial speech.
  4. 4
    Apparatus for the production of artificial speech which comprises:a source of coded representations of phonemes of speech ordered according to a desired phonetic sequence;means responsive to successions of said representations which represent discrete phonemes for generating a control signal that persists with a substantially constant magnitude for the duration of each of said phoneme representations;means responsive to successive discrete phoneme representations for generating control signals that vary in a substantially linear fashion between the pairs of said substantially constant magnitude control signals that denote vowel-to-vowel or consonantto-consonant phoneme representations;means responsive to successive discrete phoneme representations for generating control signals that vary in a substantially nonlinear fashion with a slope that monotonically increases in magnitude between pairs of said substantially constant magnitude control signals that denote a consonant-tovowel phoneme representative sequence;means responsive to successive discrete phoneme representations for generating control signals that vary in a substantially nonlinear fashion with a slope that monotonically approaches zero between pairs of said substantially constant magnitude control signals that denote a vowel-to-consonant phoneme sequence;and means for utilizing said substantially constant magnitude control signals and said varying control signals together for the generation of artificial speech.
  5. 5
    In combination, means for storing a succession of coded representations of a selected alphabet of phonemes acording to a desired phonetic order, means for storing for each phoneme a set of analog representations of parameters uniquely associated therewith, signal generator means for developing from sets of said analog representation control signals representative of substantially steady-state phoneme values, means for transferring sets of said analog representations to said signal generator means in accordance with said order of storage of said coded representations, means associated with said signal generator means for analyzing sets of analog representations applied to said generator, means responsive to analyses of successive pairs of analog representations for developing substantially nonlinear control signal segments for interconnecting respectively the control signals corresponding to said sets of analog representations, speech synthesizing means including means for generating hiss energy and buzz energy and for shaping said hiss and buzz energy spectra, and means for utilizing said control signals for controlling the shaping of said energy spectra in said synthesizer to produce intelligible speech.
  6. 6
    The combination as defined in claim 5 in further combination with means for pre-emphasizing said hiss energy at a rate of substantially 6 db per octave and for de-emphasizing hiss energy at the rate of substantially 6 db per octave, and means for adding the equalized hiss and buzz energy together to form a composite speech excitation signal.
  7. 7
    Control signal generator apparatus for producing resonance vocoder control signals in response to coded analog representations of ordered speech phonemes that comprises means for developing from said stored analog data a substantially constant signal for each discrete 3,158,685 phoneme and the next consecutive one in said order, said means including means responsive to the vowel-consonant veloping from said analog data a substantially nonlinear control signal portion for each transition between one phoneme and the next consecutive one in said order, said means including means responsive to the vowel-consonant order of successive pairs of phonemes for altering the mode of transition of said nonlinear control signals, and means for supplying one of said control signals for each control function required by speech synthesizing means. 10
  8. 8
    In combination with apparatus as defined in claim 7, means operative upon the occurrence of a stored analog representation of one of the voiceless stops, p, t, k, for interposing an abrupt discontinuity in the synthesizer con5 trol signal that relates to hiss energy. References Cited in the file of this patent UNITED STATES PATENTS 2,595,701 Potter___'_______________May 6, 1952 2,771,509 Dudley et al____________Nov. 20, 1956