EP2056578A2

Providing a multi-modal communications infrastructure for automated call centre operation

Abstract

A system (30) and method for providing a multi-modal communications infrastructure for automated call center (11) operation is provided. A multi-modal call is accepted from a caller (12-15) through a telephony interface (35), which accommodates multi-modal calls including at least one of verbal speech and text messaging. Incoming speech in the multi-modal call is converted into transcribed text (57). Incoming text messaging is matched with the transcribed text (57). The multi-modal call is automatically assigned through a session manager (31) to a session under operation of a live agent (81). The transcribed text (57) and incoming text messaging are progressively processed during the session through an agent application (43) by performing a customer support scenario (22) interactively monitored and controlled by the live agent (81).

EP2056578A2, drawing sheet 1
Sheet 1 of 21

Term

2.1 yearsto projected expiry

Projected expiry 31 October 2028, counted from filing; an application has no term until it is granted.

  1. Priority
  2. Filed
  3. Published
  4. Today
  5. Projected expiry

14 claims: 4 independent, 10 dependent

  1. 1
    A system (30) for providing a multi-modal communications infrastructure for automated call center (11) operation, comprising:a telephony interface (35) configured to accept a multi-modal call from a caller (12-15);a messaging server (31) configured to automatically assign the multi-modal call through a session manager (47) to a session under operation of a live agent (81);a speech recognition module (36) configured to convert incoming speech (54) in the multi-modal call into transcribed text (57);a text message processor configured to match incoming text messaging with the transcribed text (57);and an agent console (39) configured to progressively process the transcribed text (57) and the incoming text messaging during the session through an agent application (43) by performing a customer support scenario (22) interactively monitored and controlled by the live agent (81).
  2. 2
    A system (30) according to Claim 1, further comprising at least one of:a database (34) configured to store the incoming text messaging and the transcribed text (57) as a log entry (44);and a post-processing module configured to perform post-processing on the incoming text messaging and the transcribed text (57).
  3. 3
    A system (30) according to Claim 1, further comprising:a caller identification module configured to assign a caller identification (71) to the incoming text messaging;and the messaging server (31) configured to forward the incoming text messaging to the live agent (81) based on the caller identification (71).
  4. 4
    A system (30) according to Claim 1, further comprising:an identification (71) module configured to flag the incoming text messaging as verbatim caller data;and a database (34) configured to store the flagged incoming text messaging.
  5. 5
    A method for providing a multi-modal communications infrastructure for automated call center (11) operation, comprising:accepting a multi-modal call from a caller (12-15) through a telephony interface (35) and automatically assigning the multi-modal call through a session manager (47) to a session under operation of a live agent (81);converting incoming speech (54) in the multi-modal call into transcribed text (57);matching incoming text messaging with the transcribed text (57);and progressively processing the transcribed text (57) and the incoming text messaging during the session through an agent application (43) by performing a customer support scenario (22) interactively monitored and controlled by the live agent (81).
  6. 6
    A method according to Claim 5, further comprising at least one of:storing the incoming text messaging and the transcribed text (57) as a log entry (44);and performing post-processing on the incoming text messaging and the transcribed text (57).
  7. 7
    A method according to Claim 5, further comprising:assigning a caller identification (71) to the incoming text messaging;and forwarding the incoming text messaging to the live agent (81) based on the caller identification (71).
  8. 8
    A method according to Claim 5, further comprising:flagging the incoming text messaging as verbatim caller data;and storing the flagged incoming text messaging.
  9. 9
    A system (30) for providing service provisioning to a caller (12-15) through multi-modal communication, comprising:a message server (31) configured to assign an incoming multi-modal call from a caller (12-15) to a call session having a unique caller identification (71) and which is managed by an agent (81);a telephony interface (35) configured to receive verbal speech communication (54) from the caller (12-15) during the call session;a speech recognition module (36) configured to convert the verbal speech communication (54) to transcribed text (57) and further configured to provide the transcribed text (57) to the agent (81);a text gateway configured to receive incoming text messages from the caller (12-15);an identification (71) module to assign the unique caller identification (71) to each incoming text message and further configured to provide the incoming text messages to the agent (81) based on the unique caller identification (71);and a text-to-speech engine (37) configured to receive outgoing text messages from the agent (81) and further configured to convert the outgoing text messages into an outgoing stream of synthesized speech (63) that is presented to the caller (12-15).
  10. 10
    A system (30) according to Claim 9, further comprising:a telephony interface (35) configured to receive the verbal speech communication (54) and the incoming text messages into a call center (11) through separate streams of communication, wherein the verbal speech communication (54) and the incoming text messages simultaneously originate from the caller (12-15) and further configured to pair the verbal speech communication (54) and the incoming text messages, comprising: a caller identification (71) module configured to assign the unique caller identification (71) to the verbal speech communication (54);and a call match module configured to match the unique caller identification (71) of the verbal speech communication (54) and the incoming text messages to the call session.
  11. 11
    A system (30) according to Claim 9, further comprising:a caller information module configured to receive the multi-modal call from a mobile communication device (13), further configured to obtain caller information from the mobile communication device (13) comprising locational data (169), and further configured to analyze the locational data (169) to determine a location of the caller (12-15).
  12. 12
    A method for providing service provisioning to a caller (12-15) through multi-modal communication, comprising:assigning an incoming multi-modal call from a caller (12-15) to a call session having a unique caller identification (71) and which is managed by an agent (81);and providing service provisioning to the caller (12-15) during the multi-modal call, comprising: receiving verbal speech communication (54) from the caller (12-15) during the call session and converting the verbal speech communication (54) to transcribed text (57);providing the transcribed text (57) to the agent (81);receiving incoming text messages from the caller (12-15) and assigning the unique caller identification (71) to each incoming text message;providing the incoming text messages to the agent (81) based on the unique caller identification (71);and receiving outgoing text messages from the agent (81) and converting the outgoing text messages into an outgoing stream of synthesized speech (63) that is presented to the caller (12-15).
  13. 13
    A method according to Claim 12, further comprising:receiving the verbal speech communication (54) and the incoming text messages into a call center (11) through separate streams of communication, wherein the verbal speech communication (54) and the incoming text messages simultaneously originate from the caller (12-15);and pairing the verbal speech communication (54) and the incoming text messages, comprising: assigning the unique caller identification (71) to the verbal speech communication (54);and matching the unique caller identification (71) of the verbal speech communication (54) and the incoming text messages to the caller session.
  14. 14
    A method according to Claim 12, further comprising:receiving the multi-modal call from a mobile communication device (13);obtaining caller information from the mobile communication device (13) comprising locational data (169);and analyzing the locational data (169) to determine a location of the caller (12-15).