EP3559946A1

Facilitating end-to-end communications with automated assistants in multiple languages

Abstract

This record has no abstract on file.

Term

11.6 yearsto projected expiry

Projected expiry 16 April 2038, counted from filing; an application has no term until it is granted.

  1. Priority and filed
  2. Published
  3. Today
  4. Projected expiry

20 claims: 4 independent, 16 dependent

  1. 1
    Claims of equivalent WO 2019172946 A1 CLAIMS What is claimed is:1. A method implemented by one or more processors, comprising: receiving voice input provided by a user at an input component of a client device in a first language;generating speech recognition output from the voice input, wherein the speech recognition output is in the first language;identifying a first language intent of the user based on the speech recognition output;fulfilling the first language intent to generate first fulfillment information;based on the first fulfillment information, generating a first natural language output candidate in the first language;translating at least a portion of the speech recognition output from the first language to a second language to generate an at least partial translation of the speech recognition output;identifying a second language intent of the user based on the at least partial translation;fulfilling the second language intent to generate second fulfillment information;based on the second fulfillment information, generating a second natural language output candidate in the second language;determining scores for the first and second natural language output candidates;based on the scores, selecting, from the first and second natural language output candidates, a natural language output to be presented to the user;and causing the client device to present the selected natural language output at an output component of the client device.
  2. 9
    A method implemented by one or more processors, comprising:receiving voice input provided by a user at an input component of a client device in a first language;generating speech recognition output of the voice input in the first language;translating at least a portion of the speech recognition output from the first language to a second language to generate an at least partial translation of the speech recognition output;identifying a second language intent of the user based on the at least partial translation;fulfilling the second language intent to generate fulfillment information;generating natural language output in the second language based on the second la nguage intent;translating the natural language output to the first language to generate translated natural language output;determining whether the translated natural language output satisfies one or more criteria;based on the determining, selecting output that is based on the translated natural la nguage output or alternative natural la nguage output;and causing the client device to present the output at an output component of the client device.
  3. 13
    A system comprising one or more processors and memory operably coupled with the one or more processors, wherein the memory stores instructions that, in response to execution of the instructions by one or more processors, cause the one or more processors to perform the following operations:receiving voice input provided by a user at an input component of a client device in a first language;generating speech recognition output from the voice input, wherein the speech recognition output is in the first language;identifying a first language intent of the user based on the speech recognition output;fulfilling the first language intent to generate first fulfillment information;based on the first fulfillment information, generating a first natural language output candidate in the first language;translating at least a portion of the speech recognition output from the first language to a second language to generate an at least partial translation of the speech recognition output;identifying a second language intent of the user based on the at least partial translation;fulfilling the second language intent to generate second fulfillment information;based on the second fulfillment information, generating a second natural language output candidate in the second language;determining scores for the first and second natural language output candidates;based on the scores, selecting, from the first and second natural language output candidates, a natural language output to be presented to the user;and causing the client device to present the selected natural language output at an output component of the client device.
  4. 14
    The system of claim IB, further comprising instructions for generating a third natural language output candidate in the first language that is responsive to the second language intent, wherein determining the scores further includes determining scores for the first, second, and third content.