EP1946292A1

Method and apparatus for improving the transcription accuracy of speech recognition software

Abstract

This record has no abstract on file.

EP1946292A1, drawing sheet 1
Sheet 1 of 17

Term

Projected expiry 23 October 2026.

  1. Priority
  2. Filed
  3. Published
  4. Today
  5. Projected expiry

16 claims: 5 independent, 11 dependent

  1. 1
    Claims of equivalent WO 2007048053 A1 What is claimed is:1. A method for operating a computerized speech recognition system comprising: loading a first vocabulary;evaluating individual vocabulary elements within said first vocabulary to determine a first vocabulary match set, each vocabulary element within said first vocabulary match set having a match probability score;weighting said match probability scores of said vocabulary elements within said first vocabulary match set with a first vocabulary weighting factor;loading a second vocabulary;evaluating individual vocabulary elements within said second vocabulary to determine a second vocabulary match set, each vocabulary element within said second vocabulary match set having a match probability score;combining said individual vocabulary elements within said first and second vocabulary match sets so as to create a combine set of vocabulary elements;weighting said match probability scores of said combine set of vocabulary elements with a second vocabulary weighting factor;and selecting as a match to an input to said computerized speech recognition system a vocabulary element from said combine set of vocabulary elements based on said weighted match probability scores of said combine set of vocabulary elements.
  2. 9
    A method for operating a computerized speech recognition system comprising:loading a first vocabulary;evaluating individual vocabulary elements within said first vocabulary to determine a first vocabulary match set, each vocabulary element within said first vocabulary match set having a match probability score;loading a second vocabulary;evaluating individual vocabulary elements within said second vocabulary to determine a second vocabulary match set, each vocabulary element within said second vocabulary match set having a match probability score;combining said individual vocabulary elements within said first and second vocabulary match sets so as to create a combine set of vocabulary elements;weighting said match probability scores of said combine set of vocabulary elements with a non-linear vocabulary weighting function;evaluating individual vocabulary elements within said combined set of vocabulary elements to determine a combined vocabulary match set based on said non-linearly weighted match probability scores of said vocabulary element within combined set of vocabulary elements;and selecting as a match to an input to said computerized speech recognition system a vocabulary element from said combine set of vocabulary elements based on said weighted match probability scores of said combine set of vocabulary elements.
  3. 14
    A method for using a speech recognition system with a records database, said speech recognition system including a vocabulary database, said vocabulary database having default vocabulary elements used by said speech recognition system and imported vocabulary elements from said records database, said method comprising:providing a speech input to said speech recognition system;evaluating said speech input against said vocabulary elements within said vocabulary database to determine a probable match set of vocabulary elements, said probable match set being determined according to default weightings for said vocabulary elements as determined by said speech recognition system;providing said probable match set to a vocabulary prioritization module for use with said records database;said user prioritization module assigning use weightings to said vocabulary elements within said records database according to a plurality of use criteria and creating a predefined template according to said plurality of use criteria;creating a virtual vocabulary from said probable match set according to said predefined use template;modifying said default weights of said vocabulary elements within said probable match set according to the presence of said vocabulary elements within said virtual vocabulary;and selecting a vocabulary element as a match for said speech input based on said probable match set having said modified weightings.
  4. 15
    A method for creating a user database system for use with a speech recognition system, said user database system including a user database having user vocabulary elements, said speech recognition system including a vocabulary database having default vocabulary elements, said method comprising:importing vocabulary elements from said user database into said speech recognition system;creating a virtual vocabulary providing said probable match set to a vocabulary prioritization module for use with said user database;said user prioritization module assigning use weightings to said vocabulary elements within said user database according to a plurality of use criteria and creating a predefine use template according to said plurality of use criteria;creating a virtual vocabulary from said probable match set according to said predefined use template;modifying said default weights of said vocabulary elements within said probable match set according to the presence of said vocabulary elements within said virtual vocabulary;and selecting a vocabulary element as a match for said speech input based on said probable match set having said modified weightings.
  5. 16
    A software system for use with a computerized speech recognition system, said speech recognition system including a default vocabulary database, said default vocabulary database having vocabulary elements used by said speech recognition system, said default vocabulary elements having preexisting weightings according to a default [metric] of said speech recognition system, said system comprising:an adjunct vocabulary database having adjunct vocabulary elements, said adjunct vocabulary elements having weightings according to a [first adjunct metric], wherein said speech recognition system evaluates a speech input to said system against individual vocabulary elements within said default vocabulary according to said default weightings to create a default match set and against individual vocabulary elements within said adjunct vocabulary according to said [first adjunct weightings] to create an adjunct match set;said speech recognition system selecting a vocabulary element as a match for said speech input based on said combined default vocabulary match set and said adjunct vocabulary match set.