IP Library › Granted Patent US 12,608,170
Granted Patent B2
US 12,608,170 · App. 18/546,885 · Granted Apr 21, 2026

Interactive audio entertainment system for vehicles

Inventors: Yitshak Lior Ben Gigi (Burlington, MA); Caitlin Vachon (Burlington, MA)
Assignee: Cerence Operating Company
G06F3/165B60K35/10B60K35/26G10L15/22G10L25/54B60K2360/148G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,608,170
App. No.
18/546,885
Granted
Apr 21, 2026
Kind
B2
Abstract

A system for interacting with an audio stream to obtain lyric information, control playback of the audio stream, and control aspects of the audio stream. In some instances, end users can request that the audio stream play with a lead vocal track or without a lead vocal track. Obtaining lyric information includes receiving via a text to speech module an audio playback of the lyric information.

Claims (22)

1 . A system for interacting with an audio stream, the system comprising:

an audio playback module playing an audio stream in a first playback mode; and

a recognition module receiving, from a head unit of a vehicle, one or more utterances comprising at least one command requesting that the audio playback module play the audio stream in a second playback mode,

wherein the audio playback module responsively plays the audio stream in the second playback mode

wherein the recognition module receives a second microphone signal from the head unit, the second microphone signal including another user utterance including a second command requesting lyric information associated with the audio stream, and instructs the audio playback module to output the lyric information.

2 . The system of claim 1 , wherein the first playback mode comprises playing instrumental and lead vocal tracks of the audio stream.

3 . The system of claim 1 , wherein the second playback mode comprises playing an instrumental track but not lead vocal tracks of the audio stream.

4 . A system for interactive audio entertainment, comprising:

at least one loudspeaker configured to play back an audio stream in one or more modes into an environment;

at least one microphone configured to receive microphone signals indicative of sound in the environment; and

a processor programmed to

instruct the loudspeaker to play back the audio stream in a first playback mode,

receive a first microphone signal from the at least one microphone, the first microphone signal including a user utterance including a command to play back the audio stream in a second playback mode,

instruct the at least one loudspeaker to play back the audio stream in the second playback mode,

receive a second microphone signal from the at least one microphone, the second microphone signal including another user utterance including a second command requesting lyric information associated with the audio stream, and

instruct the at least one loudspeaker to output the lyric information.

5 . The system of claim 4 , wherein the first playback mode comprises playing instrumental and lead vocal tracks of the audio stream.

6 . The system of claim 4 , wherein the second playback mode comprises playing an instrumental track but not lead vocal tracks of the audio stream.

7 . The system of claim 4 , wherein the processor is further programmed to identify a time-bound section of the audio stream and identify the lyric information within the time-bound section of the audio stream.

8 . The system of claim 7 , wherein the time-bound section of the audio stream has a start time and a stop time.

9 . The system of claim 8 , wherein the processor is further programmed to identify the lyric information within the time-bound section of the audio stream by recognizing speech uttered between the start time and the stop time.

10 . The system of claim 8 , wherein the processor is further programmed to identify the lyric information within the time-bound section of the audio stream by searching a database for the lyric information uttered at a point in time between the start time and the stop time of the time-bound section of the audio stream.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 29, 2023
From: BEN GIGI, YITSHAK LIOR; VACHON, CAITLIN
To: CERENCE OPERATING COMPANY
Reel/Frame 064735/0141 →
Continuity (2)
Provisional Application 63151005 · Feb 18, 2021
Related Publication 20240126499A1 · Apr 18, 2024
References Cited (14)
US 5703308A · Tashiro · 1997 [cited by examiner]
US 5820384A · Tubman · 1998 [cited by examiner]
US 11245950B1 · Mahar · 2022 [cited by examiner]
US 20110273455A1 · Powar · 2011 [cited by examiner]
US 20120191461A1 · Lin · 2012 [cited by examiner]
US 20170372686A1 · Rajendran · 2017 [cited by examiner]
US 20180090116A1 · Zhao · 2018 [cited by examiner]
US 20180190307A1 · Hetherington et al. · 2018 [cited by applicant]
US 20200402490A1 · Duthaler · 2020 [cited by examiner]
US 20220375470A1 · Pledl · 2022 [cited by examiner]
JP H11167392A · 1999 [cited by examiner]
JP 2004341338A · 2004 [cited by applicant]
JP 2007047486A · 2007 [cited by applicant]
JP 2008241761A · 2008 [cited by examiner]