EP2192575B1

Speech recognition based on a multilingual acoustic model

Abstract

This record has no abstract on file.

EP2192575B1, drawing sheet 1
Sheet 1 of 10

Term

2.2 yearsleft in the term

Expires 27 November 2028.

  1. Priority and filed
  2. Granted
  3. Today
  4. Expires

9 claims: 3 independent, 6 dependent

  1. 1
    Method for generating a multilingual speech recognizer comprising a multilingual acoustic model, comprising the steps of providing a first speech recognizer comprising a first codebook consisting of first Gaussians and first Hidden Markov Models, HMMs, comprising first states;providing at least one second speech recognizer comprising a second codebook consisting of second Gaussians and second Hidden Markov Models, HMMs, comprising second states;replacing each of the second Gaussians of the at least one second speech recognizer by the respective closest one of the first Gaussians and/or each of the second states of the second HMMs of the at least one second speech recognizer with the respective closest state of the first HMMs of the first speech recognizer to obtain at least one modified second speech recognizer;and combining the first speech recognizer and the at least one modified second speech recognizer to obtain the multilingual speech recognizer.
  2. 4
    The method according to one of the preceding claims, wherein the first speech recognizer is modified by modifying the first codebook before combining it with the at least one modified second speech recognizer to obtain the multilingual speech recognizer, wherein the step of modifying the first codebook comprises adding at least one of the second Gaussians of the second codebook of the at least one second speech recognizer to the first codebook.
  3. 7
    Speech recognition means or speech dialog system or speech control system comprising a multilingual speech recognizer generated by the method according to one of the preceding claims.
  4. 8
    Audio device, in particular, an MP3 or MP4 player, cell phone or a Personal Digital Assistant, or a video device comprising a speech recognition or speech dialog system or speech control system means comprising a multilingual speech recognizer generated according to the method according to one of the claims 1 to 6.
  5. 9
    Computer program product, comprising one or more computer readable media having computer-executable instructions for performing the steps of the method according to one of the claims 1 to 6 when run on a computer.