EP3559946B1

Facilitating end-to-end communications with automated assistants in multiple languages

Abstract

This record has no abstract on file.

EP3559946B1, drawing sheet 1
Sheet 1 of 6

Term

11.6 yearsleft in the term

Expires 16 April 2038.

  1. Priority and filed
  2. Granted
  3. Today
  4. Expires

14 claims: 3 independent, 11 dependent

  1. 1
    A method implemented by one or more processors, comprising:receiving voice input provided by a user at an input component of a client device in a first language;generating speech recognition output from the voice input, wherein the speech recognition output is in the first language;identifying a first language intent of the user based on the speech recognition output;fulfilling the first language intent to generate first fulfillment information;based on the first fulfillment information, generating a first natural language output candidate in the first language;translating at least a portion of the speech recognition output from the first language to a second language to generate an at least partial translation of the speech recognition output;identifying a second language intent of the user based on the at least partial translation;fulfilling the second language intent to generate second fulfillment information;based on the second fulfillment information, generating a second natural language output candidate in the second language;determining scores for the first and second natural language output candidates;based on the scores, selecting, from the first and second natural language output candidates, a natural language output to be presented to the user;and causing the client device to present the selected natural language output at an output component of the client device.
  2. 9
    A method implemented by one or more processors, comprising:receiving voice input provided by a user at an input component of a client device in a first language;generating speech recognition output of the voice input in the first language;translating at least a portion of the speech recognition output from the first language to a second language to generate an at least partial translation of the speech recognition output;identifying a second language intent of the user based on the at least partial translation;fulfilling the second language intent to generate fulfillment information;generating natural language output in the second language based on the second language intent;translating the natural language output to the first language to generate translated natural language output;determining whether the translated natural language output satisfies one or more criteria;based on the determining, selecting output that is based on the translated natural language output or alternative natural language output;and causing the client device to present the output at an output component of the client device.
  3. 13
    A system comprising one or more processors and memory operably coupled with the one or more processors, wherein the memory stores instructions that, in response to execution of the instructions by one or more processors, cause the one or more processors to perform the method of any one of claims 1 to 12.
  4. 14
    Processor-executable instructions which, when executed by one or more processors, cause the one or more processors to perform the method of any of claims 1 to 12.