IP Library Granted Patent US 11,189,275
Granted Patent B2
US 11,189,275 · App. 16/170,327 · Granted Nov 30, 2021

Natural language processing while sound sensor is muted

Inventors: Subramanyam Irukuvajhula (Hyderabad, IN); Ravi Kiran Nalla (Hyderabad, IN)
Assignee: Polycom, Inc.
G10L15/22G06F3/165G10L15/08H04L29/06414H04M3/568G10L2015/088G10L2015/223G10L2015/228
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,189,275
App. No.
16/170,327
Granted
Nov 30, 2021
Kind
B2
Abstract

A method includes generating first audio data based on sound detected by a sound sensor at a first time. The method further includes transmitting, via one or more communication interfaces, the first audio data to another device during a communication session based on a determination that the sound sensor is unmuted with respect to the communication session at the first time. The method further includes generating second audio data based on sound detected by the sound sensor at a second time. The method further includes refraining from transmitting the second audio data to the other device during the communication session based on a determination that the sound sensor is muted with respect to the communication session at the second time. The method further includes initiating a natural language processing operation on the second audio data based on detecting a wake phrase in the second audio data.

Claims (45)

1. An apparatus comprising:

a sound sensor;

an audio output device;

one or more communication interfaces;

one or more processor devices coupled to the sound sensor, the audio output device and the one or more communication interfaces; and

one or more memory devices coupled to the one or more processor devices and storing instructions executable by the one or more processor devices to:

receive a mute data setting indicating a muted or unmuted state of a first voice communication session with a first party from a user;

generate first audio data based on sound detected by the sound sensor at a first time;

initiate transmission, via the one or more communication interfaces, of the first audio data to another device during the first voice communication session based on a determination that the sound sensor is unmuted with respect to the first voice communication session at the first time based on the state of the mute data setting indicating unmuted;

generate second audio data based on sound detected by the sound sensor at a second time;

refrain from initiating transmission of the second audio data to the other device during the first voice communication session based on a determination that the sound sensor is muted with respect to the first voice communication session at the second time based on the state of the mute data setting indicating muted;

initiate a natural language processing operation on the second audio data based on detecting a wake phrase in the second audio data, including initiating transmission, via the one or more communication interfaces, of the second audio data to a remote natural language processing device;

initiate a second voice communication session with a second party, via the one or more communication interfaces, based on a response from the natural language processing device to a spoken command detected in the second audio data;

initiate transmission, via the one or more communication interfaces, of second audio data to a device of the second party during the second voice communication session based on a determination that the sound sensor is muted with respect to the first voice communication session at the second time based on the state of the mute data setting indicating muted without placing the first voice communication session on hold;

initiate output, via the audio output device, of sound based on communication data received, via the one or more communication interfaces, from the other device during the first voice communication session and based on output data received, via the one or more communication interfaces, from the device of the second party during the second voice communication session; and

conduct the first voice communication session and the second voice communication session in parallel.

2. The apparatus of claim 1 , wherein the second voice communication session is associated with an emergency service.

3. The apparatus of claim 1 , wherein initiating the natural language processing operation includes storing, in the one or more memory devices, a note associated with the first voice communication session based on a spoken command detected in the second audio data.

4. The apparatus of claim 1 , wherein the second audio data includes the wake phrase.

5. A computer readable storage device storing instructions, the instructions executable by one or more processor devices to:

receive a mute data setting indicating a muted or unmuted state of a first voice communication session with a first party from a user;

generate first audio data based on sound detected by a sound sensor at a first time;

initiate transmission, via one or more communication interfaces, of the first audio data to another device during the first voice communication session based on a determination that the sound sensor is unmuted with respect to the first voice communication session at the first time based on the state of the mute data setting indicating unmuted;

generate second audio data based on sound detected by the sound sensor at a second time;

refrain from initiating transmission of the second audio data to the other device during the first voice communication session based on a determination that the sound sensor is muted with respect to the first voice communication session at the second time based on the state of the mute data setting indicating muted;

initiate a natural language processing operation on the second audio data based on detecting a wake phrase in the second audio data, including initiating transmission, via the one or more communication interfaces, of the second audio data to a remote natural language processing device;

initiate a second voice communication session with a second party, via the one or more communication interfaces, based on a response from the natural language processing device to a spoken command detected in the second audio data;

initiate transmission, via the one or more communication interfaces, of second audio data to a device of the second party during the second voice communication session based on a determination that the sound sensor is muted with respect to the first voice communication session at the second time based on the state of the mute data setting indicating muted without placing the first voice communication session on hold;

initiate output, via an audio output device, of sound based on communication data received, via the one or more communication interfaces, from the other device during the first voice communication session and based on output data received, via the one or more communication interfaces, from the device of the second party during the second voice communication session; and

conduct the first voice communication session and the second voice communication session in parallel.

6. The computer readable storage device of claim 5 , wherein the second voice communication session is associated with an emergency service.

7. The computer readable storage device of claim 5 , wherein initiating the natural language processing operation includes storing, in one or more memory devices, a note associated with the first voice communication session based on a spoken command detected in the second audio data.

8. The computer readable storage device of claim 5 , wherein the second audio data includes the wake phrase.

9. A method comprising:

receiving a mute data setting indicating a muted or unmuted state of a first voice communication session from a user;

generating first audio data based on sound detected by a sound sensor at a first time;

transmitting, via one or more communication interfaces, the first audio data to another device during the first voice communication session based on a determination that the sound sensor is unmuted with respect to the first voice communication session at the first time based on the state of the mute data setting indicating unmuted;

generating second audio data based on sound detected by the sound sensor at a second time;

refraining from transmitting the second audio data to the other device during the first voice communication session based on a determination that the sound sensor is muted with respect to the first voice communication session at the second time based on the state of the mute data setting indicating muted;

initiating a natural language processing operation on the second audio data based on detecting a wake phrase in the second audio data, including initiating transmission, via the one or more communication interfaces, of the second audio data to a remote natural language processing device;

initiating a second voice communication session to a second party, via the one or more communication interfaces, based on a response from the natural language processing device to a spoken command detected in the second audio data;

initiating transmission, via the one or more communication interfaces, of second audio data to a device of the second party during the second voice communication session based on a determination that the sound sensor is muted with respect to the first voice communication session at the second time based on the state of the mute data setting indicating muted without placing the first voice communication session on hold;

initiating output, via an audio output device, of sound based on communication data received, via the one or more communication interfaces, from the other device during the first voice communication session and based on output data received, via the one or more communication interfaces, from the device of the second party during the second voice communication session; and

conducting the first voice communication session and the second voice communication session in parallel.

10. The method of claim 9 , wherein the second voice communication session is associated with an emergency service.

Assignments (4)
NUNC PRO TUNC ASSIGNMENT Recorded Jun 22, 2023
From: POLYCOM, INC.
To: HEWLETT-PACKARD DEVELOPMENT COMPANY, L.P.
Reel/Frame 064056/0947 →
RELEASE OF PATENT SECURITY INTERESTS Recorded Aug 30, 2022
From: WELLS FARGO BANK, NATIONAL ASSOCIATION
To: PLANTRONICS, INC.; POLYCOM, INC.
Reel/Frame 061356/0366 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 29, 2019
From: IRUKUVAJHULA, SUBRAMANYAM; NALLA, RAVI KIRAN
To: POLYCOM, INC.
Reel/Frame 050213/0907 →
SUPPLEMENTAL SECURITY AGREEMENT Recorded Mar 6, 2019
From: PLANTRONICS, INC.; POLYCOM, INC.
To: WELLS FARGO BANK, NATIONAL ASSOCIATION
Reel/Frame 048515/0306 →
Priority Claims (1)
IN 201811029163 · Aug 2, 2018 · national
Continuity (1)
Related Publication 20200043486A1 · Feb 6, 2020