IP Library Granted Patent US 8,554,554
Granted Patent B2
US 8,554,554 · App. 13/584,006 · Granted Oct 8, 2013

Automated demographic analysis by analyzing voice activity

Inventors: Michael Johnston (New York, NY); Hisao M. Chang (Cedar Park, TX); Harry E. Blanchard (Rumson, NJ); Bernard S. Renger (New Providence, NJ); Linda Roberts (Decatur, GA)
Assignee: AT&T Intellectual Property I, L.P.
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,554,554
App. No.
13/584,006
Granted
Oct 8, 2013
Kind
B2
Abstract

Methods, systems, and media for determining a response to be generated in an environment are provided. The methods, systems, and media monitor the environment for a voice activity of an individual. The voice activity of the individual is detected and analyzed. A content descriptor of the voice activity is determined based on the voice activity of the individual. A demographic descriptor of the individual is determined based on the voice activity of the individual. The content descriptor, the demographic descriptor, and known information are correlated to determine the response to be generated in the environment.

Claims (56)

1. A method of determining a response to be generated in an environment, the method comprising:

monitoring the environment for a voice activity of an individual;

detecting the voice activity of the individual;

correlating, with a processor, the voice activity of the individual and known information of the environment, the known information corresponding to media content that is provided in the environment;

determining, with the processor and based on a correlation of the voice activity of the individual and the known information of the environment, a content descriptor of the voice activity;

determining, with the processor and based on the voice activity of the individual, a demographic descriptor of the individual; and

correlating, with the processor, the content descriptor and the demographic descriptor to determine the response to be generated in the environment.

2. The method as set forth in claim 1 , wherein the known information corresponds to visual content that is displayed in the environment.

3. The method as set forth in claim 1 , wherein the known information corresponds to audio content that is played in the environment.

4. The method as set forth in claim 1 , wherein

a plurality of voice activities is detected, each belonging to one of a plurality of individuals,

the method further comprises:

selecting one of the plurality of voice activities to be a representative voice activity of the plurality of voice activities for analyzing the representative voice activity, and

the content descriptor and the demographic descriptor are determined based on the representative voice activity.

5. The method as set forth in claim 4 , wherein the representative voice activity is selected by filtering the plurality of voice activities according to a predetermined paralinguistic property.

6. The method as set forth in claim 1 , wherein

a plurality of voice activities is detected, each belonging to one of a plurality of individuals,

the plurality of voice activities is analyzed, and

a plurality of demographic descriptors is determined for the plurality of individuals based on the plurality of voice activities.

7. The method as set forth in claim 6 , wherein the content descriptor and the plurality of demographic descriptors are correlated to determine a plurality of responses to be generated in the environment.

8. The method as set forth in claim 1 , wherein the environment is a public environment.

9. The method as set forth in claim 1 , further comprising:

detecting visual activity of the individual;

analyzing, with the processor, the visual activity of the individual; and

correlating, with the processor, the voice activity and the visual activity of the individual to determine the demographic descriptor.

10. The method as set forth in claim 1 , wherein the voice activity is detected with an audio capturer that includes an input that enables the individual to input an input activity for determining the demographic descriptor.

11. The method as set forth in claim 1 , further comprising:

storing the demographic descriptor in a database in association with the individual for creating a profile of the individual.

12. The method as set forth in claim 11 , wherein the profile associated with the individual is configured to be accessible by the individual for displaying the profile on a social networking website.

13. The method as set forth in claim 1 , further comprising:

detecting the voice activity of the individual for a predetermined period of time;

wherein the voice activity is analyzed over the predetermined period of time for determining the content descriptor and the demographic descriptor.

14. The method as set forth in claim 1 , wherein

the voice activity is detected with an audio capturer, and

the voice activity is transmitted over a network for analyzing the voice activity for determining the content descriptor and the demographic descriptor.

15. The method as set forth in claim 1 , further comprising:

receiving an input from the individual for one of accepting and rejecting the demographic descriptor.

16. A system for determining a response to be generated in an environment, the system comprising:

a processor that detects a voice activity of an individual; and

a memory storing instructions that, when executed by the processor, cause the processor to perform operations including:

correlating the voice activity of the individual and known information of the environment, the known information corresponding to media content that is provided in the environment;

determining, based on a correlation of the voice activity of the individual and the known information of the environment, a content descriptor of the voice activity;

determining, based on the voice activity of the individual, a demographic descriptor of the individual; and

correlating the content descriptor and the demographic descriptor to determine the response to be generated in the environment.

17. The system as set forth in claim 16 , wherein the known information corresponds to visual content that is displayed in the environment.

18. The system as set forth in claim 16 , wherein the known information corresponds to audio content that is played in the environment.

19. The system as set forth in claim 16 , wherein

the processor detects a plurality of voice activities, each belonging to one of a plurality of individuals, and

the processor selects one of the plurality of voice activities to be a representative voice activity of the plurality of voice activities, analyzes the representative voice activity, and determines the content descriptor and the demographic descriptor based on the representative voice activity.

20. A non-transitory computer readable medium having an executable computer program for determining a response to be generated in an environment that, when executed by a processor, causes the processor to perform operations comprising:

monitoring the environment for a voice activity of an individual;

detecting the voice activity of the individual;

correlating the voice activity of the individual and known information of the environment, the known information corresponding to media content that is provided in the environment;

determining, based on a correlation of the voice activity of the individual and the known information of the environment, a content descriptor of the voice activity;

determining, based on the voice activity of the individual, a demographic descriptor of the individual; and

correlating the content descriptor and the demographic descriptor to determine the response to be generated in the environment.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 13, 2012
From: JOHNSTON, MICHAEL; CHANG, HISAO M.; BLANCHARD, HARRY E.; RENGER, BERNARD S.; ROBERTS, LINDA
To: AT&T INTELLECTUAL PROPERTY I, L.P.
Reel/Frame 028775/0502 →
Continuity (2)
Continuation 12344981 · Dec 29, 2008
Related Publication 20120323566A1 · Dec 20, 2012