IP Library Granted Patent US 8,553,065
Granted Patent B2
US 8,553,065 · App. 13/088,974 · Granted Oct 8, 2013

System and method for providing augmented data in a network environment

Inventors: Satish K. Gannu (San Jose, CA); Leon A. Frazier (Cary, NC); Didier R. Moretti (Los Altos, CA)
Assignee: Cisco Technology, Inc.
H04N7/15
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,553,065
App. No.
13/088,974
Granted
Oct 8, 2013
Kind
B2
Abstract

A method is provided in one example and includes identifying a particular word recited by an active speaker in a conference involving a plurality of endpoints in a network environment; evaluating a profile associated with the active speaker in order to identify contextual information associated with the particular word; and providing augmented data associated with the particular word to at least some of the plurality of endpoints. In more specific examples, the active speaker is identified using a facial detection protocol, or a speech recognition protocol. Data from the active speaker can be converted from speech to text.

Claims (35)

1. A method, comprising:

Identifying, by a media engine, a particular word recited by an active speaker in a conference involving a plurality of endpoints in a network environment;

evaluating, by the media engine, a profile associated with the active speaker in order to identify contextual information associated with the particular word;

determining, by the media engine, augmented data associated with the particular word based on the contextual information identified from the profile associated with the active speaker;

providing, by the media engine, the augmented data associated with the particular word to at least some of the plurality of endpoints.

2. The method of claim 1 , wherein the active speaker is identified using a facial detection protocol, or a speech recognition protocol.

3. The method of claim 1 , wherein audio data from the active speaker is converted from speech to text.

4. The method of claim 1 , wherein information in the profile is used to provide personal data about the active speaker to the plurality of endpoints.

5. The method of claim 1 , wherein the augmented data is provided on a display of a video conference.

6. The method of claim 1 , wherein the augmented data is provided as a message in a chat forum of an audio conference.

7. The method of claim 1 , wherein the augmented data is provided as audio data delivered to at least some of the endpoints on particular end-user devices.

8. The method of claim 1 , wherein profiles associated with the plurality of endpoints are evaluated in order to adjust the augmented data to be provided to each of the endpoints.

9. Logic encoded in non-transitory media that includes code for execution and when executed by a processor is operable to perform operations comprising:

identifying a particular word recited by an active speaker in a conference involving a plurality of endpoints in a network environment;

evaluating a profile associated with the active speaker in order to identify contextual information associated with the particular word;

determining augmented data associated with the particular word based on the contextual information identified from the profile associated with the active speaker;

providing augmented data associated with the particular word to at least some of the plurality of endpoints.

10. The logic of claim 9 , wherein the active speaker is identified using a facial detection protocol, or a speech recognition protocol.

11. The logic of claim 9 , wherein audio data from the active speaker is converted from speech to text.

12. The logic of claim 9 , wherein information in the profile is used to provide personal data about the active speaker to the plurality of endpoints.

13. The logic of claim 9 , wherein the augmented data is provided on a display of a video conference.

14. The logic of claim 9 , wherein the augmented data is provided as a message in a chat forum of an audio conference.

15. The logic of claim 9 , wherein the augmented data is provided as audio data delivered to at least some of the endpoints on particular end-user devices.

16. An apparatus, comprising:

a memory element configured to store data;

a processor operable to execute instructions associated with the data;

a media engine configured to interface with the memory element and the processor such that the apparatus is configured for:

identifying a particular word recited by an active speaker in a conference involving a plurality of endpoints in a network environment;

evaluating a profile associated with the active speaker in order to identify contextual information associated with the particular word;

determining augmented data associated with the particular word based on the contextual information identified from the profile associated with the active speaker; and

providing augmented data associated with the particular word to at least some of the plurality of endpoints.

17. The apparatus of claim 16 , wherein the active speaker is identified using a facial detection protocol, or a speech recognition protocol.

18. The apparatus of claim 16 , wherein audio data from the active speaker is converted from speech to text.

19. The apparatus of claim 16 , wherein information in the profile is used to provide personal data about the active speaker to the plurality of endpoints.

20. The apparatus of claim 16 , wherein profiles associated with the plurality of endpoints are evaluated in order to adjust the augmented data to be provided to each of the endpoints.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 18, 2011
From: GANNU, SATISH K.; FRAZIER, LEON A.; MORETTI, DIDIER R.
To: CISCO TECHNOLOGY, INC.
Reel/Frame 026144/0777 →
Continuity (1)
Related Publication 20120262533A1 · Oct 18, 2012