US20160111084A1

Speech recognition device and speech recognition method

Claim Score by NHIP

Read claim 8, the broadest

Abstract

A speech recognition device includes: a collector collecting speech data of a first speaker from a speech-based device; a first storage accumulating the speech data of the first speaker; a learner learning the speech data of the first speaker accumulated in the first storage and generating an individual acoustic model of the first speaker based on the learned speech data; a second storage storing the individual acoustic model of the first speaker and a generic acoustic model; a feature vector extractor extracting a feature vector from the speech data of the first speaker when a speech recognition request is received from the first speaker; and a speech recognizer selecting either one of the individual acoustic model of the first speaker and the generic acoustic model based on an accumulated amount of the speech data of the first speaker and recognizing a speech command using the extracted feature vector and the selected acoustic model.

US20160111084A1, drawing sheet 1
Sheet 1 of 4

Term

Projected expiry 28 July 2035.

  1. Priority
  2. Filed
  3. Published
  4. Today
  5. Projected expiry

15 claims: 3 independent, 12 dependent

  1. 1
    A speech recognition device comprising:a collector collecting speech data of a first speaker from a speech-based device;a first storage accumulating the speech data of the first speaker;a learner learning the speech data of the first speaker accumulated in the first storage and generating an individual acoustic model of the first speaker based on the learned speech data;a second storage storing the individual acoustic model of the first speaker and a generic acoustic model;a feature vector extractor extracting a feature vector from the speech data of the first speaker when a speech recognition request is received from the first speaker;and a speech recognizer selecting either one of the individual acoustic model of the first speaker and the generic acoustic model based on an accumulated amount of the speech data of the first speaker and recognizing a speech command using the extracted feature vector and the selected acoustic model.
  2. 8
    Broadest claimClaim Score 58, broad(NHIP)A speech recognition method comprising:collecting speech data of a first speaker from a speech-based device;accumulating the speech data of the first speaker in a first storage;learning the accumulated speech data of the first speaker;generating an individual acoustic model of the first speaker based on the learned speech data;storing the individual acoustic model of the first speaker and a generic acoustic model in a second storage;extracting a feature vector from the speech data of the first speaker when a speech recognition request is received from the first speaker;selecting either one of the individual acoustic model of the first speaker and the generic acoustic model based on an accumulated amount of the speech data of the first speaker;and recognizing a speech command using the extracted feature vector and the selected acoustic model.
  3. 15
    A non-transitory computer readable medium containing program instructions for performing a speech recognition method, the computer readable medium comprising:program instructions that collect speech data of a first speaker from a speech-based device;program instructions that accumulate the speech data of the first speaker in a first storage;program instructions that learn the accumulated speech data of the first speaker;program instructions that generate an individual acoustic model of the first speaker based on the learned speech data;program instructions that store the individual acoustic model of the first speaker and a generic acoustic model in a second storage;program instructions that extract a feature vector from the speech data of the first speaker if when a speech recognition request is received from the first speaker;program instructions that select either one of the individual acoustic model of the first speaker and the generic acoustic model based on an accumulated amount of the speech data of the first speaker;and program instructions that recognize a speech command using the extracted feature vector and the selected acoustic model.