IP Library Granted Patent US 9,361,887
Granted Patent B1
US 9,361,887 · App. 14/846,925 · Granted Jun 7, 2016

System and method for providing words or phrases to be uttered by members of a crowd and processing the utterances in crowd-sourced campaigns to facilitate speech analysis

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,361,887
App. No.
14/846,925
Granted
Jun 7, 2016
Kind
B1
Abstract

Systems and methods of providing text related to utterances, and gathering voice data in response to the text are provide herein. In various implementations, an identification token that identifies a first file for a voice data collection campaign, and a second file for a session script may be received from a natural language processing training device. The first file and the second file may be used to configure the mobile application to display a sequence of screens, each of the sequence of screens containing text of at least one utterance specified in the voice data collection campaign. Voice data may be received from the natural language processing training device in response to user interaction with the text of the at least one utterance. The voice data and the text may be stored in a transcription library.

Claims (41)

1. A computer-implemented method, the method being implemented in a computer system having one or more physical processors programmed with computer program instructions that, when executed by the one or more physical processors, cause the computer system to perform the method, the method comprising:

receiving from a natural language processing training device an identification token containing a first portion and a second portion, the first portion identifying a first file for a voice data collection campaign, and the second portion identifying a second file for a session script, the session script supporting a mobile application on the natural language processing training device;

using the first file and the second file to configure the mobile application to display a sequence of screens, each of the sequence of screens containing text of at least one utterance specified in the voice data collection campaign;

receiving voice data from the natural language processing training device in response to user interaction with the text of the at least one utterance; and

storing the voice data and the text of the at least one utterance in a transcription library.

2. The method of claim 1 , further comprising:

gathering a first filename corresponding to the first file;

gathering a second filename corresponding to the second file; and

creating the identification token using the first filename and the second filename.

3. The method of claim 1 , wherein the identification token comprises an alphanumeric character string.

4. The method of claim 1 , wherein the identification token comprises a concatenation of the first portion and the second portion.

5. The method of claim 1 , wherein one or more of the first file and the second file comprises a JavaScript Object Notation (JSON) file.

6. The method of claim 1 , wherein the utterance comprises one or more of a syllable, a word, a phrase, or a variant thereof.

7. The method of claim 1 , wherein the user interaction comprises a selection of a touch-screen button instructing the mobile application to record the voice data.

8. The method of claim 1 , wherein the natural language processing training device comprises one or more of a mobile phone, a tablet computing device, a laptop, and a desktop.

9. The method of claim 1 , wherein the voice data collection campaign is configured to collect demographic information related to a user of the natural language processing training device.

10. A system comprising:

a memory;

one or more physical processors programmed with one or more computer program instructions which, when executed, cause the one or more physical processors to:

receive from a natural language processing training device an identification token containing a first portion and a second portion, the first portion identifying a first file for a voice data collection campaign, and the second portion identifying a second file for a session script, the session script supporting a mobile application on the natural language processing training device;

use the first file and the second file to configure the mobile application to display a sequence of screens, each of the sequence of screens containing text of at least one utterance specified in the voice data collection campaign;

receive voice data from the natural language processing training device in response to user interaction with the text of the at least one utterance; and

store the voice data and the text of the at least one utterance in a transcription library.

11. The system of claim 10 , wherein the instructions cause the one or more physical processors to:

gather a first filename corresponding to the first file;

gather a second filename corresponding to the second file; and

create the identification token using the first filename and the second filename.

12. The system of claim 10 , wherein the identification token comprises an alphanumeric character string.

13. The system of claim 10 , wherein the identification token comprises a concatenation of the first portion and the second portion.

14. The system of claim 10 , wherein one or more of the first file and the second file comprises a JavaScript Object Notation (JSON) file.

15. The system of claim 10 , wherein the utterance comprises one or more of a syllable, a word, a phrase, or a variant thereof.

16. The system of claim 10 , wherein the user interaction comprises a selection of a touch-screen button instructing the mobile application to record the voice data.

17. The system of claim 10 , wherein the natural language processing training device comprises one or more of a mobile phone, a tablet computing device, a laptop, and a desktop.

18. The system of claim 10 , wherein the voice data collection campaign is configured to collect demographic information related to a user of the natural language processing training device.

19. A computer program product comprising:

one or more tangible, non-transitory computer-readable storage devices;

program instructions, stored on at least one of the one or more tangible, non-transitory computer-readable tangible storage devices that, when executed, cause a computer to:

receive from a natural language processing training device an identification token containing a first portion and a second portion, the first portion identifying a first file for a voice data collection campaign, and the second portion identifying a second file for a session script, the session script supporting a mobile application on the natural language processing training device;

use the first file and the second file to configure the mobile application to display a sequence of screens, each of the sequence of screens containing text of at least one utterance specified in the voice data collection campaign;

receive voice data from the natural language processing training device in response to user interaction with the text of the at least one utterance; and

store the voice data and the text of the at least one utterance in a transcription library.

Assignments (8)
RELEASE (REEL 052935 / FRAME 0584) Recorded Jan 2, 2025
From: WELLS FARGO BANK, NATIONAL ASSOCIATION
To: CERENCE OPERATING COMPANY
Reel/Frame 069797/0818 →
SECURITY AGREEMENT Recorded Jun 15, 2020
From: CERENCE OPERATING COMPANY
To: WELLS FARGO BANK, N.A.
Reel/Frame 052935/0584 →
RELEASE OF SECURITY INTEREST Recorded Jun 12, 2020
From: BARCLAYS BANK PLC
To: CERENCE OPERATING COMPANY
Reel/Frame 052927/0335 →
SECURITY AGREEMENT Recorded Nov 7, 2019
From: CERENCE OPERATING COMPANY
To: BARCLAYS BANK PLC
Reel/Frame 050953/0133 →
INTELLECTUAL PROPERTY AGREEMENT Recorded Oct 24, 2019
From: NUANCE COMMUNICATIONS, INC.
To: CERENCE INC.
Reel/Frame 050818/0001 →
RELEASE OF SECURITY INTEREST Recorded Apr 5, 2018
From: ORIX GROWTH CAPITAL, LLC
To: VOICEBOX TECHNOLOGIES CORPORATION
Reel/Frame 045581/0630 →
SECURITY INTEREST Recorded Dec 22, 2017
From: VOICEBOX TECHNOLOGIES CORPORATION
To: ORIX GROWTH CAPITAL, LLC
Reel/Frame 044949/0948 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 29, 2015
From: BRAGA, DANIELA; ROMANI, FARAZ; ELSHENAWY, AHMAD KHAMIS; KENNEWICK, MICHAEL
To: VOICEBOX TECHNOLOGIES CORPORATION
Reel/Frame 036914/0348 →