IP Library › Granted Patent US 8,706,473
Granted Patent B2
US 8,706,473 · App. 13/231,635 · Granted Apr 22, 2014

System and method for insertion and removal of video objects

Inventors: Jim Chen Chou (San Jose, CA); Rohit Puri (Campbell, CA); Tapabrata Biswas (Sunnyvale, CA)
Assignee: Cisco Technology, Inc.
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,706,473
App. No.
13/231,635
Granted
Apr 22, 2014
Kind
B2
Abstract

An example method may include receiving a media stream from a first endpoint, where the media stream is intended for a second endpoint; processing the media stream according to at least one processing option; compressing the media stream; and communicating the media stream to the second endpoint. In more specific instances, the processing may include converting a speech in the media stream to text in a first language; converting the text in the first language to text in a second language; rendering the text in the second language; and adding the rendered text in the media stream.

Claims (66)

1. A method, comprising:

receiving a video stream from a first endpoint;

video processing the video stream according to an Internet Protocol (IP) address of a second endpoint, wherein the video processing the video stream comprises

extracting a network time protocol (NTP) timestamp from the video stream,

decoding the NTP timestamp to a local time, and

adding the local time as text in the video stream; and

communicating the video stream to the second endpoint.

2. The method of claim 1 , further comprising:

converting a speech in a media stream including the video stream to text in a first language;

converting the text in the first language to text in a second language;

rendering the text in the second language; and

adding the rendered text in the media stream.

3. The method of claim 1 , further comprising:

compressing the video stream.

4. The method of claim 1 , wherein the video processing the video stream comprises:

inserting or removing an advertisement in the video stream.

5. The method of claim 4 , wherein the advertisement comprises text associated with a multipoint video conference.

6. The method of claim 1 , wherein the video stream is video processed according to a selected one of a group of processing options, the group consisting of:

a device profile associated with at least one of the endpoints;

a geographical location associated with at least one of the endpoints; and

a language preference associated with at least one of the endpoints.

7. Logic encoded in non-transitory media that includes code for execution and, when executed by a processor, is operable to perform operations comprising:

receiving a video stream from a first endpoint;

video processing the video stream according to an Internet Protocol (IP) address of a second endpoint, wherein the video processing the video stream comprises

extracting a network time protocol (NTP) timestamp from the video stream,

decoding the NTP timestamp to a local time, and

adding the local time as text in the video stream; and

communicating the video stream to the second endpoint.

8. The logic of claim 7 , wherein the operations further comprise:

converting a speech in a media stream including the video stream to text in a first language;

converting the text in the first language to text in a second language;

rendering the text in the second language; and

adding the rendered text in the media stream.

9. The logic of claim 7 , wherein the video processing the video stream comprises:

removing video objects by using an image segmentation protocol in conjunction with at least one motion estimation calculation.

10. The logic of claim 7 , wherein the video processing the video stream comprises:

inserting or removing an advertisement in the video stream.

11. The logic of claim 10 , wherein the advertisement comprises text associated with a multipoint video conference.

12. The logic of claim 7 , wherein the video stream is video processed according to a selected one of a group of processing options, the group consisting of:

a device profile associated with at least one of the endpoints;

a geographical location associated with at least one of the endpoints; and

a language preference associated with at least one of the endpoints.

13. An apparatus, comprising:

a processor operable to execute instructions;

a memory; and

a control module configured to interface with the processor such that

the apparatus is configured to

receive a video stream from a first endpoint;

video process the video stream according to an Internet Protocol (IP) address of a second endpoint; and

communicate the video stream to the second endpoint, and

the processor is configured to

extract a network time protocol (NTP) timestamp from the video stream;

decode the NTP timestamp to a local time; and

add the local time as text in the video stream.

14. The apparatus of claim 13 , wherein the processor is configured to

convert a speech in a media stream including the video stream to text in a first language;

convert the text in the first language to text in a second language;

render the text in the second language; and

add the rendered text in the media stream.

15. The apparatus of claim 13 , wherein the processor is configured to insert or remove an advertisement in the video stream.

16. The apparatus of claim 15 , wherein the advertisement comprises text associated with a multipoint video conference.

17. The apparatus of claim 13 , wherein the video stream is video processed according to a selected one of a group of processing options, the group consisting of:

a device profile associated with at least one of the endpoints;

a geographical location associated with at least one of the endpoints; and

a language preference associated with at least one of the endpoints.

18. The apparatus of claim 13 , wherein the apparatus is located at an edge of the network, which includes at least one video endpoint.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 13, 2011
From: CHOU, JIM CHEN; PURI, ROHIT; BISWAS, TAPABRATA
To: CISCO TECHNOLOGY, INC.
Reel/Frame 026897/0708 →
Continuity (1)
Related Publication 20130066623A1 · Mar 14, 2013