IP Library Granted Patent US 6,894,715
Granted Patent B2
US 6,894,715 · App. 09/883,475 · Granted May 17, 2005

Mixing video signals for an audio and video multimedia conference call

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 6,894,715
App. No.
09/883,475
Granted
May 17, 2005
Kind
B2
Abstract

In a multimedia communications system ( 100 ) that supports conference calls that include an audio portion and a video portion, a primary video image is selected from a plurality of video images based on an amount of audio data generated. The amount of audio data is determined by counting a number of audio packets or by counting an amount of audio samples in audio packets ( 204 ). A dominant audio participant is selected if the difference in the amount of audio exceeds a predetermined threshold ( 206 ). If the difference in the amount of audio does not exceed the predetermined threshold ( 206 ), the dominant audio participant may be determined by comparing the loudness or volume for each audio participant ( 207 212 ). The primary video image is selected to correspond to the dominant audio participant ( 208, 214 ). The primary video image remains constant for a predetermined period of time before the possibility to change ( 210, 216 ).

Claims (34)

1. In a communications system that supports conference calls that include an audio portion and a video portion, a method for selecting a primary video image from a plurality of video images, the method comprising the steps of:

receiving audio data in a digital form;

determining an amount of the audio data in digital form generated by each participant of a plurality of participants in a conference call;

selecting a dominating audio participant from the plurality of participants based upon the amount of the audio data in digital form generated by each participant of the plurality of participants;

and selecting a primary video image based on the dominating audio participant,

wherein the step of determining an amount of the audio data in digital form comprises counting a number of one of audio packets and audio samples in packet form generated by each participant of the plurality of participants.

2. The method of claim 1 wherein the step of determining an amount of the audio data in digital form comprises counting an amount of audio samples in audio packets.

3. The method of claim 1 wherein the primary video image is larger than a plurality of remaining video images of the plurality of video images.

4. The method of claim 1 further comprising the step of maintaining the primary video image for at least a predetermined period of time.

5. In a communications system that supports conference calls that include an audio portion and a video portion, a method for selecting a primary video image from a plurality of video images, the method comprising the steps of:

receiving audio data in a digital form;

determining an amount of the audio data in digital form generated by each participant of a plurality of participants in a conference call;

determining whether a difference between an amount of the audio data in digital form generated by one participant of the plurality of participants and an amount of the audio data in digital form generated by other participants of the plurality of participants exceeds a predetermined threshold;

if the difference exceeds the predetermined threshold, then selecting a dominating audio participant from the plurality of participants based upon the amount of the audio data in digital form generated by each participant of the plurality of participants; and

selecting a primary video image based on the dominating audio participant,

wherein the step of determining an amount of the audio data in digital form comprises counting a number of one of the audio samples in packet form generated by each participant of the plurality of participants.

6. The method of claim 5 wherein the dominating audio participant generates an amount of the audio data in digital form that exceeds an amount of the audio data in digital form generated by each of a plurality of remaining participants of the plurality of participants.

7. The method of claim 5 further comprising the step of:

if the difference does not exceed the predetermined threshold, then determining a loudness of audio for each participant of the plurality of participants; and

selecting the dominating audio participant based on the loudness for each participant of the plurality of participants.

8. The method of claim 5 wherein the step of determining an amount of the audio data in digital form comprises counting an amount of audio samples in audio packets.

9. The method of claim 5 wherein the primary video image is larger than a plurality of remaining video images of the plurality of video images.

10. The method of claim 5 further comprising the step of maintaining the primary video image for at least a predetermined period of time.

11. In a communications system that supports conference calls that include an audio portion and a video portion, an apparatus for selecting a primary video image from a plurality of video images, the apparatus comprising:

a first processor that:

receives audio data in a digital form; and

determines an amount of the audio data in digital form generated by each participant of a plurality of participants in a conference call;

a second processor that selects a dominating audio participant from the plurality of participants based upon the amount of the audio data in digital form generated by each participant of the plurality of participants; and

a third processor that selects a primary video image based on the dominating audio participant,

wherein the first processor determines an amount of the audio data in digital form by counting a number of one of audio packets and audio samples in packet form generated by each participant of the plurality of participants.

12. The apparatus of claim 11 wherein the first processor, the second processor and the third processor are a same processor.

13. The apparatus of claim 11 wherein at least two of the first processor, the second processor and the third processor are a same processor.

14. The apparatus of claim 11 wherein the primary video image is larger than a plurality of remaining video images of the plurality of video images.

15. The apparatus of claim 11 wherein the first processor determines an amount of the audio data by counting an amount of the audio samples in audio packets.

Assignments (7)
NUNC PRO TUNC ASSIGNMENT Recorded Oct 8, 2019
From: NOKIA OF AMERICA CORPORATION
To: ALCATEL LUCENT
Reel/Frame 050668/0829 →
CHANGE OF NAME Recorded Sep 24, 2019
From: ALCATEL-LUCENT USA INC.
To: NOKIA OF AMERICA CORPORATION
Reel/Frame 050476/0085 →
RELEASE OF SECURITY INTEREST Recorded Oct 9, 2014
From: CREDIT SUISSE AG
To: ALCATEL-LUCENT USA INC.
Reel/Frame 033950/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 25, 2014
From: ALCATEL LUCENT
To: SOUND VIEW INNOVATIONS, LLC
Reel/Frame 033416/0763 →
MERGER Recorded Sep 30, 2013
From: LUCENT TECHNOLOGIES INC.
To: ALCATEL-LUCENT USA INC.
Reel/Frame 031309/0403 →
SECURITY INTEREST Recorded Mar 7, 2013
From: ALCATEL-LUCENT USA INC.
To: CREDIT SUISSE AG
Reel/Frame 030510/0627 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 16, 2001
From: HENRIKSON, ERIC HAROLD
To: LUCENT TECHNOLOGIES INC.
Reel/Frame 011943/0580 →