EP1083545A2

Voice recognition of proper names in a navigation apparatus

Abstract

A voice recognition apparatus includes: a voice input device; a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition; and a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through the voice input device and voice recognition data created in correspondence to the recognition word, and the storage device stores both a first recognition word corresponding to a pronunciation of an entirety of the word to undergo voice recognition and a second recognition word corresponding to a pronunciation of only a starting portion of a predetermined length of the entirety of the word to undergo voice recognition as recognition words for the word to undergo voice recognition.

EP1083545A2, drawing sheet 1
Sheet 1 of 26

Term

Term ended

Projected expiry passed 7 September 2020, 6 years ago.

  1. Priority
  2. Filed
  3. Published
  4. Projected expiry
  5. Today

30 claims: 19 independent, 11 dependent

  1. 1
    A voice recognition apparatus, comprising:a voice input device;a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition;and a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data created in correspondence to said recognition word, wherein: said storage device stores both a first recognition word corresponding to a pronunciation of an entirety of said word to undergo voice recognition and a second recognition word corresponding to a pronunciation of only a starting portion of a predetermined length of the entirety of said word to undergo voice recognition as recognition words for said word to undergo voice recognition.
  2. 3
    A voice recognition navigation apparatus, comprising:a voice input device;a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition;a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data created in correspondence to said recognition word;a map information storage device that stores map information;and a control device that engages in control for providing route guidance based upon, at least, recognition results obtained by said voice recognition processing device and said map information, wherein: said storage device stores both a first recognition word corresponding to a pronunciation of an entirety of said word to undergo voice recognition and a second recognition word corresponding to a pronunciation of only a starting portion of a predetermined length of the entirety of said word to undergo voice recognition as recognition words for said word to undergo voice recognition.
  3. 4
    A voice recognition apparatus, comprising:a voice input device;a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition;and a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data created in correspondence to said recognition word, wherein: said storage device stores a plurality of recognition words each having a different pronunciation, for a single word to undergo voice recognition.
  4. 5
    A voice recognition apparatus, comprising:a voice input device;a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition;and a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data created in correspondence to said recognition word, wherein: said storage device stores both a first recognition word corresponding to the pronunciation of an entirety of said word to undergo voice recognition and a second recognition word created by replacing the leading syllable in the pronunciation of the entirety of said word to undergo voice recognition with a vowel constituting the leading syllable, as recognition words for the word to undergo voice recognition.
  5. 6
    A voice recognition apparatus, comprising:a voice input device;a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition;and a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data created in correspondence to said recognition word, wherein: said storage device stores both a first recognition word corresponding to the pronunciation of an entirety of said word to undergo voice recognition and a second recognition word created by deleting a starting portion of a predetermined length of the pronunciation of the entirety of said word to undergo voice recognition, as recognition words for the word to undergo voice recognition.
  6. 9
    A voice recognition navigation apparatus, comprising:a voice input device;a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition;a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data created in correspondence to said recognition word;a map information storage device that stores map information;and a control device that engages in control for providing route guidance based upon, at least, recognition results obtained by said voice recognition processing device and said map information, wherein: said storage device stores a plurality of recognition words each having a different pronunciation, for a single word to undergo voice recognition.
  7. 10
    A voice recognition apparatus, comprising:a voice input device;a storage device that stores recognition words to be used in voice recognition processing;and a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data generated based upon said recognition words, wherein: said storage device stores valid recognition words corresponding to pronunciations of words to undergo voice recognition and invalid recognition words each indicating a pronunciation that is dissimilar to the pronunciations of said words to undergo voice recognition;and when said audio data obtained through said voice input device manifests a highest similarity to voice recognition data generated based upon one of said invalid recognition words, said voice recognition processing device decides that none of the words to undergo voice recognition has been recognized.
  8. 12
    A voice recognition apparatus, comprising:a voice input device;a storage device that stores recognition words to be used in voice recognition processing;a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data generated based upon said recognition words;a map information storage device that stores map information, and a control device that engages in control for providing route guidance based upon, at least, voice recognition results obtained by said voice recognition processing device and said map information, wherein: said storage device stores valid recognition words corresponding to pronunciations of words to undergo voice recognition and invalid recognition words each indicating a pronunciation that is dissimilar to the pronunciations of said words to undergo voice recognition;and when said audio data obtained through said voice input device manifests a highest similarity to voice recognition data generated based upon one of said invalid recognition words, said voice recognition processing device decides that none of the words to undergo voice recognition has been recognized.
  9. 13
    A voice recognition apparatus, comprising:a voice input device;a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition;and a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data created in correspondence to said recognition word, wherein: when a given word to undergo voice recognition includes a predetermined specific word as a part of the given word, a first recognition word created by replacing a standard pronunciation of said specific word with an alternative pronunciation of said specific word different from said standard pronunciation is stored in said storage device.
  10. 19
    A voice recognition apparatus, comprising:a voice input device;a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition;and a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data created in correspondence to said recognition word, wherein: if a predetermined specific word is not included in said word to undergo voice recognition, a recognition word created by adding a pronunciation of said specific word is stored in said storage device stores.
  11. 20
    A voice recognition navigation apparatus, comprising:a voice input device;a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition;and a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data created in correspondence to said recognition word;a map information storage device that stores map information;and a control device that engages in control for providing route guidance based upon, at least, recognition results obtained by said voice recognition processing device and said map information, wherein: when a given word to undergo voice recognition includes a predetermined specific word as a part of the given word, a first recognition word created by replacing a standard pronunciation of said specific word with an alternative pronunciation different from said standard pronunciation is stored in said storage device.
  12. 21
    A voice recognition navigation apparatus, comprising:a voice input device;a storage device that stores a recognition word indicating a pronunciation of a word to undergo voice recognition;and a voice recognition processing device that performs voice recognition processing by comparing audio data obtained through said voice input device and voice recognition data created in correspondence to said recognition word;a map information storage device that stores map information;and a control device that engages in control for providing route guidance based upon, at least, recognition results obtained by said voice recognition processing device and said map information, wherein: if a predetermined specific word is not included in said word to undergo voice recognition, a recognition word created by adding a pronunciation of said specific word is stored in said storage device.
  13. 22
    A method of recognition word generation through which recognition words indicating pronunciations of words to undergo voice recognition used to generate voice recognition data to be compared against audio data obtained through a voice input device are generated, comprising;a step in which when a given word to undergo voice recognition contains a predetermined specific word as a part of the given word, a recognition word is created by replacing a standard pronunciation of said specific word with a alternative pronunciation different from said standard pronunciation.
  14. 23
    Data representing recognition words corresponding to a word to undergo voice recognition that is used to generate voice recognition data to be compared against audio data obtained through a voice input device in voice recognition processing, characterized by comprising:a first recognition word corresponding to a pronunciation of an entirety of said word to undergo voice recognition;and a second recognition word corresponding to a pronunciation of only a starting portion of a predetermined length of the entirety of said word to undergo voice recognition, wherein both said first recognition word and said second recognition word are used as recognition words for said word to undergo voice recognition.
  15. 24
    Data representing recognition words corresponding to a word to undergo voice recognition that is used to generate voice recognition data to be compared against audio data obtained through a voice input device in voice recognition processing, characterized by comprising:a first recognition word corresponding to a pronunciation of an entirety of said word to undergo voice recognition;and a second recognition word created by replacing a leading syllable in the pronunciation of the entirety of the word to undergo voice recognition with a vowel constituting said leading syllable, wherein both said first recognition word and said second recognition word are used as recognition words for said word to undergo voice recognition.
  16. 25
    Data representing recognition words corresponding to a word to undergo voice recognition that is used to generate voice recognition data to be compared against audio data obtained through a voice input device in voice recognition processing, characterized by comprising:a first recognition word corresponding to a pronunciation of an entirety of said word to undergo voice recognition;and a second recognition word created by deleting a starting portion of a predetermined length of the pronunciation of the entirety of the word to undergo voice recognition, wherein both said first recognition word and said second recognition word are used as recognition words for the word to undergo voice recognition.
  17. 26
    A voice recognition control program characterized by comprising:an instruction in which audio data generated based upon a voice that has been input are compared with voice recognition data generated based upon valid recognition words corresponding to words to undergo voice recognition and indicating pronunciations of the words or invalid recognition words each indicating a pronunciation dissimilar to the pronunciations of all said words to undergo voice recognition;and an instruction in which it is decided that none of the words to undergo voice recognition has been recognized if the audio data manifest a highest similarity to voice recognition data generated based upon one of said invalid recognition words as comparison results.
  18. 27
    A recognition word generating program for generating recognition words indicating pronunciations of words to undergo voice recognition used to generate voice recognition data to be compared against audio data obtained through a voice input device in voice recognition processing, characterized by comprising:an instruction in which, if a given word to undergo voice recognition includes a predetermined specific word as a part of the given word, a recognition word is generated by replacing a standard pronunciation of said specific word with an alternative pronunciation of said specific word different from said standard pronunciation.
  19. 28
    Data representing recognition words indicating pronunciations of words to undergo voice recognition used to create voice recognition data to be compared against audio data obtained through a voice input device in voice recognition processing, characterized by comprising:when a given word to undergo voice recognition includes a predetermined specific word as a part the given word, a recognition word created by replacing a standard pronunciation of said specific word with an alternative pronunciation of said specific word different from said standard pronunciation
Independent claims19