SYSTEMS AND METHODS TO AUTOMATICALLY PERFORM ACTIONS BASED ON MEDIA CONTENT
Systems and methods are provided for automatically performing an action in respect of a conference call. One example method includes receiving, at a computing device, audio and determining a user response to the audio. Audio content is determined with natural language processing. An action based on the user response and the audio content is performed.
1 . A method for automatically performing an action in respect of a conference call, the method comprising:
receiving, at a computing device, audio;
determining a user response to the audio;
determining, with natural language processing, audio content; and
performing an action based on the user response and the audio content.
2 . The method of claim 1 , wherein the computing device further comprises an image capture device and wherein:
the method further comprises capturing one or more images of the user via the image capture device; and
determining a user response to the audio further comprises identifying, based on the one or more captured images, the user response.
3 . The method of claim 1 , wherein the computing device further comprises an image capture device and wherein:
the method further comprises capturing one or more images of the user via the image capture device; and
determining a user response to the audio further comprises:
determining, based on the one or more images, a facial expression of the user; and
identifying, based on the facial expression, the user response.
4 . The method of claim 1 , wherein the computing device further comprises an image capture device and wherein:
the method further comprises capturing one or more images of the user via the image capture device; and
determining a user response to the audio further comprises:
determining, based on the one or more images, an emotion of the user; and
identifying, based on the emotion, the user response.
5 . The method of claim 1 , wherein the computing device further comprises an audio capture device and wherein:
receiving audio comprises capturing audio of the user via the audio capture device; and
determining a user response to the audio further comprises identifying, based on the captured audio, the user response.
6 . The method of claim 1 , wherein the computing device further comprises an audio capture device and wherein:
receiving audio comprises capturing audio of the user via the audio capture device; and
determining a user response to the audio further comprises:
identifying, based on the captured audio, a characteristic associated with the user's voice; and
identifying, based on the characteristic, the user response.
7 . The method of claim 1 , wherein determining a user response to the audio further comprises monitoring the time a conferencing program is displayed on a display of the computing device.
8 . The method of claim 1 , wherein the computing device further comprises a display and an eye tracking device and wherein:
the method further comprises identifying a portion of the display that the user focuses on via the eye tracking device; and
determining a user response to the audio further comprises identifying, based on the identified portion of the display, the user response.
9 . The method of claim 1 , wherein the method further comprises:
identifying a user interest profile, the user interest profile comprising an association between audio content and a user response; and
predicting, based on the user profile, a user response to received audio content.
10 . The method of claim 1 , wherein the action comprises alerting the user to specific audio content.
11 . A system for automatically performing an action in respect of a conference call, the system comprising:
a communication port; and
control circuitry configured to:
receive, at a computing device, audio;
determine a user response to the audio;
determine, with natural language processing, audio content; and
perform an action based on the user response and the audio content.
12 . The system of claim 11 , wherein:
the control circuitry is further configured to capture one or more images of the user via a image capture device of the computing device; and
the control circuity configured to determine a user response to the audio is further configured to identify, based on the one or more captured images, the user response.
13 . The system of claim 11 , wherein:
the control circuitry is further configured to capture one or more images of the user via a image capture device of the computing device; and
the control circuity configured to determine a user response to the audio is further configured to:
determine, based on the one or more images, a facial expression of the user; and
identify, based on the facial expression, the user response.
14 . The system of claim 11 , wherein:
the control circuitry is further configured to capture one or more images of the user via a image capture device of the computing device; and
the control circuity configured to determine a user response to the audio is further configured to:
determine, based on the one or more images, an emotion of the user; and
identify, based on the emotion, the user response.
15 . The system of claim 11 , wherein:
the control circuitry configured to receive audio is further configured to capture audio of the user via an audio capture device of the computing device; and
the control circuity configured to determine a user response to the audio is further configured to identify, based on the captured audio, the user response.
16 . The system of claim 11 , wherein:
the control circuity configured to receive audio is further configured to capture audio of the user via an audio capture device of the computing device; and
the control circuitry configured to determine a user response to the audio is further configured to:
determine a characteristic associated with the user's voice; and
identify, based on the characteristic, the user response.
17 . The system of claim 11 , wherein the control circuity configured to determine a user response to the audio is further configured to monitor the time a conferencing program is displayed on a display of the computing device.
18 . The system of claim 11 , wherein:
the control circuity is further configured to identify a portion of the display that the user focuses on via an eye tracking device of the computing device; and
the control circuity configured to determine a user response to the audio is further configured to identify, based on the identified portion of the display, the user response.
19 . The system of claim 11 , wherein the control circuity is further configured to:
identify a user interest profile, the user interest profile comprising an association between audio content and a user response; and
predict, based on the user interest profile, a user response to received audio content.
20 . The system of claim 11 , wherein the control circuitry configured to perform an action is further configured to alert the user to specific audio content.
21 .- 30 . (canceled)