US20180032611A1

Systems and methods for automatic-generation of soundtracks for live speech audio

Claim Score by NHIP

Read claim 18, the broadest

Abstract

A method of automatically generating a digital soundtrack for playback in an environment comprising live speech audio generated by one or more persons speaking in the environment, the method executed by a processing device or devices having associated memory. The method comprises syntactically and/or semantically analysing an incoming text data stream or streams representing or corresponding to the live speech audio in portions to generate an emotional profile for each text portion of the text data stream(s) in the context of a continuous emotion model. The method further comprises generating in real-time a customised soundtrack for the live speech audio comprising music tracks that are played back in the environment in real-time with the live speech audio. Each music track is selected for playback in the soundtrack based at least partly on the determined emotional profile or profiles associated with the most recently processed portion or portions of text from the text data stream(s).

US20180032611A1, drawing sheet 1
Sheet 1 of 15

Term

Projected expiry 28 July 2037.

  1. Priority
  2. Filed
  3. Published
  4. Today
  5. Projected expiry

18 claims: 3 independent, 15 dependent

  1. 1
    A method of automatically generating a digital soundtrack for playback in an environment comprising live speech audio generated by one or more persons speaking in the environment, the method executed by a processing device or devices having associated memory, the method comprising:generating or receiving or retrieving an incoming live speech audio stream or streams representing the live speech audio into memory for processing;generating or retrieving or receiving an incoming text data stream or streams representing or corresponding to the live speech audio stream(s), the text data corresponding to the spoken words in the live speech audio streams;continuously or periodically or arbitrarily applying semantic processing to a portion or portions of text from the incoming text data stream(s) to determine an emotional profile associated with the processed portion or portions text;and generating in real-time a customised soundtrack comprising at least music tracks that are played back in the environment in real-time with the live speech audio, and wherein the method comprises selecting each music track for playback in the soundtrack based at least partly on the determined emotional profile or profiles associated with the most recently processed portion or portions of text from the text data stream(s).
  2. 17
    A method of automatically generating a digital soundtrack for playback in an environment comprising live speech audio generated by one or more persons speaking in the environment, the method executed by a processing device or devices having associated memory, the method comprising:receiving or retrieving an incoming live speech audio stream representing the live speech audio in memory for processing in portions;generating or retrieving or receiving text data representing or corresponding to the speech audio of each portion or portions of the incoming audio stream in memory;syntactically and/or semantically analysing the current and subsequent portions of text data in memory in the context of a continuous emotion model to generate respective emotional profiles for each of the current and subsequent portions of incoming text data;and continuously generating a soundtrack for playback in the environment that comprises dynamically selected music tracks for playback, each new music track cued for playback being selected based at least partly on the generated emotional profile associated with the most recently analysed portion of text data in memory.
  3. 18
    Broadest claimClaim Score 47, average(NHIP)A method of automatically generating a digital soundtrack on demand for playback with in an environment comprising live speech audio generated by one or more persons speaking in the environment, the method executed by a processing device or devices having associated memory, the method comprising:receiving or retrieving an incoming speech audio stream representing the live speech audio;generating or retrieving or receiving a stream of text data representing or corresponding to the incoming speech audio stream;processing the stream of text data in portions by syntactically and/or semantically analysing each portion of text data in the context of a continuous emotion model to generate respective emotional profiles for each portion of text data;and continuously generating a soundtrack for playback in the environment by selecting and co-ordinating music tracks for playback based on processing the generated emotional profiles the portions of text data.