IP Library Granted Patent US 9,564,123
Granted Patent B1
US 9,564,123 · App. 14/704,833 · Granted Feb 7, 2017

Method and system for building an integrated user profile

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,564,123
App. No.
14/704,833
Granted
Feb 7, 2017
Kind
B1
Abstract

A system and method are provided for adding user characterization information to a user profile by analyzing user's speech. User properties such as age, gender, accent, and English proficiency may be inferred by extracting and deriving features from user speech, without the user having to configure such information manually. A feature extraction module that receives audio signals as input extracts acoustic, phonetic, textual, linguistic, and semantic features. The module may be a system component independent of any particular vertical application or may be embedded in an application that accepts voice input and performs natural language understanding. A profile generation module receives the features extracted by the feature extraction module and uses classifiers to determine user property values based on the extracted and derived features and store these values in a user profile. The resulting profile variables may be globally available to other applications.

Claims (83)

1. A method for assigning values to user profile properties based on user speech analysis, comprising:

receiving extracted features from a feature extraction module, including at least acoustic features, the extracted features representing user speech observed over a first time period;

deriving features by executing a function that uses one or more of the extracted features as input, wherein at least one of the derived features is derived from an extracted acoustic feature;

storing, using a processor, derived features;

computing, using a processor, extended features as a function of one or more of:

the extracted features;

the derived features; and

statistics based on the extracted features and the derived features received during a second time period that is longer than the first time period;

deriving a value of a user profile property by executing a classifier that uses the computed extended features as input, the classifier trained on labeled extended features for the user profile property; and

storing, by a processor, the value assigned to the user profile property in a user profile.

2. The method of claim 1 , wherein the classifier is trained using one or more of:

a pattern recognition technique; and

a machine learning technique.

3. The method of claim 1 , wherein the user profile property is one of:

age;

gender;

accent;

reading level;

education level;

English language proficiency; and

socio-economic status (SES).

4. The method of claim 1 , wherein a second value of a second user profile property is used as a constraint on the value of the user profile property.

5. The method of claim 1 , wherein extracted features are received from a plurality of feature extractors.

6. The method of claim 1 , wherein the extracted features include a plurality of: a cepstrum, a spectrogram, and a phoneme lattice.

7. A system for assigning values to user profile properties based on user speech analysis, comprising:

a processor; and

a non-transitory computer-readable storage medium storing instructions, which when executed by the processor, cause the processor to:

receive extracted features, including at least acoustic features, the extracted features representing user speech observed over a first time period;

derive features by executing a function that uses one or more of the extracted features as input, wherein at least one of the derived features is derived from an extracted acoustic feature;

store derived features;

compute extended features as a function of one or more of:

the extracted features;

the derived features; and

statistics based on the extracted features and the derived features received during a second time period that is longer than the first time period;

derive a value of a user profile property by executing a classifier that uses the computed extended features as input, the classifier trained on labeled extended features for the user profile property; and

store the value assigned to the profile property in a user profile.

8. The system of claim 7 , wherein the classifier is trained using one or more of:

a pattern recognition technique; and

a machine learning technique.

9. The system of claim 7 , wherein the user profile property is one of:

age;

gender;

accent;

reading level;

education level;

English language proficiency; and

socio-economic status (SES).

10. The system of claim 7 , wherein a second value of a second user profile property is used as a constraint on the value of the user profile property.

11. The system of claim 7 , wherein extracted features are received from a plurality of feature extractors.

12. A non-transitory computer-readable storage medium storing instructions for assigning values to user profile properties based on user speech analysis, the instructions, which when executed by a processor, cause the processor to:

receive extracted features, including at least acoustic features, the extracted features representing user speech observed over a first time period;

derive features by executing a function that uses one or more of the extracted features as input, wherein at least one of the derived features is derived from an extracted acoustic feature;

store derived features;

compute extended features as a function of one or more of:

the extracted features;

the derived features; and

statistics based on the extracted features and the derived features received during a second time period that is longer than the first time period;

derive the value of a user profile property by executing a classifier that uses the computed extended features as input, the classifier trained on labeled extended features for the user profile property; and

store the value assigned to the profile property in a user profile.

13. The non-transitory computer-readable storage medium of claim 12 , wherein the classifier is trained using one or more of:

a pattern recognition technique; and

a machine learning technique.

14. The non-transitory computer-readable storage medium of claim 12 , wherein the user profile property is one of:

age;

gender;

accent;

reading level;

education level;

English language proficiency; and

socio-economic status (SES).

15. The non-transitory computer-readable storage medium of claim 12 , wherein a second value of a second user profile property is used as a constraint on the value of the user profile property.

16. The non-transitory computer-readable storage medium of claim 12 , wherein extracted features are received from a plurality of feature extractors.

17. A method to determine user profile property values based on textual input, comprising:

receiving from a text-based natural language application a textual input entered by a user;

creating a parse tree by processing the textual input against a grammar;

deriving linguistic features from the parse tree;

computing, using a processor, extended features as a function of one or more of:

the linguistic features; and

statistics based on the derived linguistic features;

deriving the value of a user profile property by executing a classifier that uses the computed extended features as input, the classifier trained on labeled extended features for the user profile property; and

storing the value assigned to the user profile property in a user profile.

18. The method of claim 17 , wherein the derived linguistic features include an indication of use of grammar or vocabulary.

19. The method of claim 1 , wherein the received extracted features, representing user speech observed over a first time period further includes at least one of: phonetic features, textual features, linguistic features, and semantic features.

Assignments (12)
TERMINATION AND RELEASE OF SECURITY INTEREST IN PATENTS Recorded Dec 3, 2024
From: MONROE CAPITAL MANAGEMENT ADVISORS, LLC, AS COLLATERAL AGENT
To: SOUNDHOUND, INC.
Reel/Frame 069480/0312 →
SECURITY INTEREST Recorded Aug 9, 2024
From: SOUNDHOUND, INC.
To: MONROE CAPITAL MANAGEMENT ADVISORS, LLC, AS COLLATERAL AGENT
Reel/Frame 068526/0413 →
RELEASE OF SECURITY INTEREST Recorded Jun 11, 2024
From: ACP POST OAK CREDIT II LLC, AS COLLATERAL AGENT
To: SOUNDHOUND, INC.; SOUNDHOUND AI IP, LLC
Reel/Frame 067698/0845 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 27, 2023
From: SOUNDHOUND AI IP HOLDING, LLC
To: SOUNDHOUND AI IP, LLC
Reel/Frame 064205/0676 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 23, 2023
From: SOUNDHOUND, INC.
To: SOUNDHOUND AI IP HOLDING, LLC
Reel/Frame 064083/0484 →
RELEASE OF SECURITY INTEREST Recorded Apr 21, 2023
From: FIRST-CITIZENS BANK & TRUST COMPANY, AS AGENT
To: SOUNDHOUND, INC.
Reel/Frame 063411/0396 →
RELEASE OF SECURITY INTEREST Recorded Apr 19, 2023
From: OCEAN II PLO LLC, AS ADMINISTRATIVE AGENT AND COLLATERAL AGENT
To: SOUNDHOUND, INC.
Reel/Frame 063380/0625 →
SECURITY INTEREST Recorded Apr 17, 2023
From: SOUNDHOUND, INC.; SOUNDHOUND AI IP, LLC
To: ACP POST OAK CREDIT II LLC
Reel/Frame 063349/0355 →
CORRECTIVE ASSIGNMENT TO CORRECT THE COVER SHEET PREVIOUSLY RECORDED AT REEL: 056627 FRAME: 0772. ASSIGNOR(S) HEREBY CONFIRMS THE SECURITY INTEREST. Recorded Apr 12, 2023
From: SOUNDHOUND, INC.
To: OCEAN II PLO LLC, AS ADMINISTRATIVE AGENT AND COLLATERAL AGENT
Reel/Frame 063336/0146 →
SECURITY INTEREST Recorded Jun 18, 2021
From: OCEAN II PLO LLC, AS ADMINISTRATIVE AGENT AND COLLATERAL AGENT
To: SOUNDHOUND, INC.
Reel/Frame 056627/0772 →
SECURITY INTEREST Recorded Apr 1, 2021
From: SOUNDHOUND, INC.
To: SILICON VALLEY BANK
Reel/Frame 055807/0539 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 18, 2015
From: MONT-REYNAUD, BERNARD; HUANG, JUN; LOKESWARAPPA, KIRAN GARAGA; GEDALIUS, JOEL
To: SOUNDHOUND, INC.
Reel/Frame 035663/0204 →