CA2806180C

Efficiently reducing transcription error using hybrid voice transcription

Abstract

A system (10) and method (40) for efficiently reducing transcription error using hybrid voice transcription is provided. A voice stream (22) is parsed into utterances. An initial transcribed value and corresponding recognition score (113) are assigned to each utterance. A transcribed message (23) is generated and includes the initial transcribed values. A threshold is applied to the recognition scores (113) to identify those utterances with recognition scores (113) below the threshold as questionable utterances (111). At least one questionable utterance (111) is compared to other questionable utterances (111) from other calls and a group of similar questionable utterances (111) is formed. One or more of the similar questionable utterances (111) is selected from the group. A common manual transcription value (124) is received for the selected similar questionable utterances (111). The common manual transcription value (124) is assigned to the remaining similar questionable utterances (111) in the group.

CA2806180C, drawing sheet 1
Sheet 1 of 9

Term

4.8 yearsleft in the term

Expires 13 July 2031.

  1. Priority
  2. Filed
  3. Granted
  4. Today
  5. Expires

20 claims: 7 independent, 13 dependent

  1. 1
    CLAIMS:1. A system Tor efficiently reducing transcription error via hybrid voice transcription, comprising: a parser to parse a voice stream from each call in a group of calls into utterances and to assign an initial transcribed value and corresponding recognition score to each utterance within the voice streams;a message generator to generate a transcribed message for each call comprising the initial transcribed values for the voice stream associated with that call;a threshold module to identify those utterances with recognition scores below a threshold as questionable utterances;a comparison module to compare at least one of the questionable utterances from the call to other questionable utterances from other calls in the group and to determine within a predetermined time frame a group of the at least one questionable utterance and other questionable utterances that are similar to the at least one questionable utterance;and an assignment module to receive a common manual transcription value for at least a portion of the similar questionable utterances and to assign the common manual transcription value to non-selected similar questionable utterances remaining in the group.
  2. 2
    A system according to Claim 1, further comprising:a replacement module to replace in the transcribed message the initial transcribed value for the at least one questionable utterance with the common manual transcription value.
  3. 3
    A system according to Claim I, further comprising:a pool generator to build the group of similar questionable utterances based on at least one of a size threshold and a time threshold.
  4. 4
    A system according to Claim I, further comprising;a sample module to select the portion of similar questionable utterances based on one of random sampling, specific sampling, and a combination of random and specific sampling. CA 2806180 2018-09-27
  5. 5
    A system according to Claim I, further comprising:a similarity module to determine a similarity between the at least one questionable utterance and the other questionable utterances based on at least one of the initial transcribed value, recognition score, range of similarity, and characteristics of the call.
  6. 6
    A method for efficiently reducing transcription error via hybrid voice transcription, comprising:parsing a voice stream from each call in a group of calls into utterances and assigning an initial transcribed value and corresponding recognition score to each utterance within the voice streams;generatings transcribed message for each call comprising the initial transcribed values for the voice stream associated with that call;identifying those utterances with recognition scores below a threshold as questionable utterances;comparing at least one questionable utterance from the call to other questionable utterances from other calls in the group;determining within a predetermined time frame a group of other questionable utterances that are similar to the at least one questionable utterance and the at least one questionable utterance;receiving a common manual transcription value for at least a portion of the similar questionable utterances;and assigning the common manual transcription value to non-selected similar questionable utterances remaining in the group.
  7. 7
    A method according to Claim 6, further comprising:replacing in the transcribed message the initial transcribed value for the at least one questionable utterance with the common manual transcription value.
  8. 8
    A method according to Claim 6, further comprising:building the group of similar questionable utterances based on at least one of a size threshold and a time threshold. CA 2806180 2018-09-27
  9. 9
    A method according to Claim 6, further comprising:selecting the portion of similar questionable utterances based on one of random sampling, specific sampling, and a combination of random and specific sampling.
  10. 10
    A method according to Claim 6. further comprising:determining a similarity between the at least one questionable utterance and the other questionable utterances based on at least one of the initial transcribed value, recognition score, range of similarity, and characteristics of the call.
  11. 11
    A system for hybrid voice transcription error reduction, comprising:a parser to parse calls in a group into speech utterances;a message generator to generate a transcribed message for each of the calls by assigning an initial transcribed value and corresponding recognition score to each utterance in that call;a threshold module to apply a confidence threshold to the recognition scores and to select those utterances with initial transcribed values that fall below the threshold as questionable utterances;a sample module to generate a sample for at least one of the questionable utterances comprising determining within a predetermined time frame for identification of similar questionable utterances other questionable utterances from other calls in the group that are assigned similar initial transcribed values as the similar questionable utterances, grouping the at least one questionable utterance with the similar questionable utterances as related utterances, and selecting a portion of the related utterances in the group as the sample: and a module to receive a common manual transcription value for each of the selected related utterances and to assign the common manual transcription value of each of the selected related utterances to each non-selected related utterance remaining in the group. CA 2806180 2018-09-27
  12. 12
    A system according to Claim 11, further comprising:a pool generator to form the group of the related utterances by identifying a threshold number of the related utterances within the predetermined time limit.
  13. 13
    A system according to Claim 11, further comprising:a replacement module to replace in the transcribed message the initial transcribed value with the common manual transcription value for the at least one questionable utterance.
  14. 14
    A system according to Claim 11, wherein the sample is selected using at least one of random sampling, specific selection sampling, and a combination of random and specific selection sampling.
  15. 15
    A system according to Claim 11, further comprising:a similarity module to determine the similarity of the al least one questionable utterance and the other questionable utterances based on one or more of the initial transcribed value, recognition score, range of similarity between the related utterances, and characteristics of the call.
  16. 16
    A method for hybrid voice transcription error reduction, comprising:parsing calls in a group into speech utterances and generating a transcribed message for each of the calls by assigning an initial transcribed value and corresponding recognition score to each utterance in that call;applying a confidence threshold to the recognition scores and selecting those utterances with initial Iranscrilied values that fall below the threshold as questionable utterances;generating a sample for at least one of the questionable utterances, comprising: determining within a predetermined time frame for identification of similar questionable utterances other questionable utterances from other calls in the group that are assigned similar initial transcribed values as the similar questionable utterances;CA 2806180 2018-09-27 grouping the at least one questionable utterance with the similar questionable utterances as related utterances;and selecting a portion of the related utterances as the sample;and receiving a common manual transcription value for each of the related utterances and assigning the common manual transcription value of each of the selected related utterances to non-selected related utterance remaining in the group.
  17. 17
    A method according to Claim 16. further comprising:forming the group of the related utterances by identifying a threshold number of the related utterances within the predetermined time limit.
  18. 18
    A method according to Claim 16, further comprising:replacing in the transcribed message the initial transcribed value with the common manual transcription value for the at least one questionable utterance.
  19. 19
    A method according to Claim 16, wherein the sample is selected using at least one of random sampling, specific selection sampling, and a combination of random and specific selection sampling.
  20. 20
    A method according to Claim 16. further comprising:determining the similarity of the at least one questionable utterance and the other questionable utterances based on one or more of the initial transcribed value, recognition score, range of similarity between the related utterances, and characteristics of the call.
Independent claims20