System and method for providing augmented data in a network environment
A method is provided in one example and includes identifying a particular word recited by an active speaker in a conference involving a plurality of endpoints in a network environment; evaluating a profile associated with the active speaker in order to identify contextual information associated with the particular word; and providing augmented data associated with the particular word to at least some of the plurality of endpoints. In more specific examples, the active speaker is identified using a facial detection protocol, or a speech recognition protocol. Data from the active speaker can be converted from speech to text.
1. A method, comprising:
Identifying, by a media engine, a particular word recited by an active speaker in a conference involving a plurality of endpoints in a network environment;
evaluating, by the media engine, a profile associated with the active speaker in order to identify contextual information associated with the particular word;
determining, by the media engine, augmented data associated with the particular word based on the contextual information identified from the profile associated with the active speaker;
providing, by the media engine, the augmented data associated with the particular word to at least some of the plurality of endpoints.
2. The method of claim 1 , wherein the active speaker is identified using a facial detection protocol, or a speech recognition protocol.
3. The method of claim 1 , wherein audio data from the active speaker is converted from speech to text.
4. The method of claim 1 , wherein information in the profile is used to provide personal data about the active speaker to the plurality of endpoints.
5. The method of claim 1 , wherein the augmented data is provided on a display of a video conference.
6. The method of claim 1 , wherein the augmented data is provided as a message in a chat forum of an audio conference.
7. The method of claim 1 , wherein the augmented data is provided as audio data delivered to at least some of the endpoints on particular end-user devices.
8. The method of claim 1 , wherein profiles associated with the plurality of endpoints are evaluated in order to adjust the augmented data to be provided to each of the endpoints.
9. Logic encoded in non-transitory media that includes code for execution and when executed by a processor is operable to perform operations comprising:
identifying a particular word recited by an active speaker in a conference involving a plurality of endpoints in a network environment;
evaluating a profile associated with the active speaker in order to identify contextual information associated with the particular word;
determining augmented data associated with the particular word based on the contextual information identified from the profile associated with the active speaker;
providing augmented data associated with the particular word to at least some of the plurality of endpoints.
10. The logic of claim 9 , wherein the active speaker is identified using a facial detection protocol, or a speech recognition protocol.
11. The logic of claim 9 , wherein audio data from the active speaker is converted from speech to text.
12. The logic of claim 9 , wherein information in the profile is used to provide personal data about the active speaker to the plurality of endpoints.
13. The logic of claim 9 , wherein the augmented data is provided on a display of a video conference.
14. The logic of claim 9 , wherein the augmented data is provided as a message in a chat forum of an audio conference.
15. The logic of claim 9 , wherein the augmented data is provided as audio data delivered to at least some of the endpoints on particular end-user devices.
16. An apparatus, comprising:
a memory element configured to store data;
a processor operable to execute instructions associated with the data;
a media engine configured to interface with the memory element and the processor such that the apparatus is configured for:
identifying a particular word recited by an active speaker in a conference involving a plurality of endpoints in a network environment;
evaluating a profile associated with the active speaker in order to identify contextual information associated with the particular word;
determining augmented data associated with the particular word based on the contextual information identified from the profile associated with the active speaker; and
providing augmented data associated with the particular word to at least some of the plurality of endpoints.
17. The apparatus of claim 16 , wherein the active speaker is identified using a facial detection protocol, or a speech recognition protocol.
18. The apparatus of claim 16 , wherein audio data from the active speaker is converted from speech to text.
19. The apparatus of claim 16 , wherein information in the profile is used to provide personal data about the active speaker to the plurality of endpoints.
20. The apparatus of claim 16 , wherein profiles associated with the plurality of endpoints are evaluated in order to adjust the augmented data to be provided to each of the endpoints.