IP Library Granted Patent US 9,369,666
Granted Patent B2
US 9,369,666 · App. 14/424,683 · Granted Jun 14, 2016

Video conference systems implementing orchestration models

Inventors: Emmanuel Marilly (Nozay, FR); Alaeddine Mihoub (Nozay, FR); Abdelkader Outtagarts (Nozay, FR)
Assignee: Alcatel Lucent
H04N7/147H04L65/1093H04L65/4038H04N7/152
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,369,666
App. No.
14/424,683
Granted
Jun 14, 2016
Kind
B2
Abstract

A method for generating an output video stream in a video conference comprising receiving a plurality of input video streams of the video conference, receiving a series of observation events ( 52, 53, 54 ), the observation corresponding to actions made by participants of the video conference, Providing a plurality of orchestration models, Determining, for each of the orchestration models a probability of the series of observation events received, Selecting an orchestration model corresponding to the highest probability, Using the selected orchestration model to perform the steps of: • selecting the display state ( 51, 40, 41, 42 ) as a candidate display state, • Determining a conditional probability of the candidate display state for the received series of observation events • Determining the candidate display state providing the highest conditional probability as an updated display state, • Generating a video stream comprising the current display state and the updated display state.

Claims (55)

1. A method for generating an output video stream in a video conference comprising:

Receiving a plurality of input video streams of the video conference

Receiving a series of observation events, the observation events belonging to a plurality of observable actions corresponding to actions made by participants of the video conference,

Providing a plurality of orchestration models, each model comprising:

A set of display states, each one associated with a predefined screen template, each screen template comprising a selected subset of the input video streams,

Transition probabilities between the display states,

Observation probabilities representing the conditional probabilities of the observable actions as a function of the display states,

Determining, for each of the orchestration models a probability of the series of observation events received,

Selecting an orchestration model corresponding to the highest probability

Using the selected orchestration model to perform:

For each display state of the orchestration model, selecting the display state as a candidate display state,

Determining a conditional probability of the candidate display state for the received series of observation events taking into account a sequence of display states including past display states and a current display state,

Determining the candidate display state providing the highest conditional probability as an updated display state,

Generating a video stream comprising one after the other a first sequence of images representing the screen template associated to the current display state and a second sequence of images representing the screen template associated to the updated display state.

2. A method according to claim 1 , wherein the observable actions are selected in the group of action categories of gestures, head motions, face expressions, audio actions, enunciation of keywords, actions relating to presentation slides.

3. A method according to claim 1 , wherein the observable actions are selected in the group of:

raising a finger, raising a hand,

making a head top down movement, making a head right left movement,

making a face expression that corresponds to speaking or sleeping,

making a noise, making silence, speaking by the tutor, speaking by a participant,

enunciating a name of an auditor or a subtitle,

switching a slide, moving a pointer,

beginning a question, ending a question.

4. A method in accordance with claim 1 , wherein the input video streams are selected in a group of: views of individual participants, views of a speaker, views of a conference room and views of presentation slides.

5. A method in accordance with claim 1 , wherein a screen template comprises a predefined arrangement of the input video streams belonging to the corresponding subset.

6. A method in accordance with claim 1 , wherein the transition probabilities are arranged as a transition matrix.

7. A method in accordance with claim 1 , wherein observation probabilities are arranged as an emission matrix.

8. A video conference control device for generating an output video stream in a video conference, the device comprising:

Means for receiving a plurality of input video streams of the video conference,

Means for receiving a series of observation events, the observation events belonging to a plurality of observable actions corresponding to actions made by participants of the video conference,

A data repository storing a plurality of orchestration models, each model comprising:

A set of display states, each one associated with a predefined screen template, each screen template comprising a selected subset of the input video streams,

Transition probabilities between the display states,

Observation probabilities representing the conditional probabilities of the observable actions as a function of the display states,

Means for determining, for each of the orchestration models, a probability of the series of observation events received,

Means for selecting an orchestration model corresponding to the highest probability,

Means for using the selected orchestration model to perform the steps of:

For each display state of the orchestration model, selecting the display state as a candidate display state,

Determining a conditional probability of the candidate display state for the received series of observation events taking into account a sequence of display states including past display states and a current display state,

Determining the candidate display state providing the highest conditional probability as an updated display state,

Generating a video stream comprising one after the other a first sequence of images representing the screen template associated to the current display state and a second sequence of images representing the screen template associated to the updated display state.

9. A video conference control device according to claim 8 , wherein the observable actions are selected in the group of action categories of gestures, head motions, face expressions, audio actions, enunciation of keywords, actions relating to presentation slides.

10. A video conference control device in accordance with claim 8 , wherein the observable actions are selected in the group of:

raising a finger, raising a hand,

making a head top down movement, making a head right left movement,

making a face expression that corresponds to speaking or sleeping,

making a noise, making silence, speaking by the tutor, speaking by a participant,

enunciating a name of an auditor or a subtitle,

switching a slide, moving a pointer,

beginning a question, ending a question.

11. A video conference control device in accordance with claim 8 , wherein the input video streams are selected in a group of: views of individual participants, views of a speaker, views of a conference room and views of presentation slides.

12. A video conference control device in accordance with claim 8 , wherein a screen template comprises a predefined arrangement of the input video streams belonging to the corresponding subset.

13. A video conference control device in accordance with claim 8 , wherein the transition probabilities are arranged as a transition matrix.

14. A video conference control device in accordance with claim 8 , wherein observation probabilities are arranged as an emission matrix.

15. A video conference system comprising a video conference control device in accordance with claim 8 , connected by a communication network to a plurality of terminals, wherein each terminal comprises means for generating an input video stream and wherein the communication network is adapted to transmit the video stream from the terminals to the control device and to transmit the output video stream generated by the control device to a terminal.

Assignments (11)
PATENT SECURITY AGREEMENT Recorded Aug 6, 2024
From: RPX CORPORATION; RPX CLEARINGHOUSE LLC
To: BARINGS FINANCE LLC, AS COLLATERAL AGENT
Reel/Frame 068328/0674 →
RELEASE OF LIEN ON PATENTS Recorded Aug 5, 2024
From: BARINGS FINANCE LLC
To: RPX CORPORATION
Reel/Frame 068328/0278 →
PATENT SECURITY AGREEMENT Recorded Apr 22, 2023
From: RPX CORPORATION
To: BARINGS FINANCE LLC, AS COLLATERAL AGENT
Reel/Frame 063429/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 28, 2021
From: PROVENANCE ASSET GROUP LLC
To: RPX CORPORATION
Reel/Frame 059352/0001 →
RELEASE OF SECURITY INTEREST Recorded Nov 30, 2021
From: NOKIA US HOLDINGS INC.
To: PROVENANCE ASSET GROUP HOLDINGS LLC; PROVENANCE ASSET GROUP LLC
Reel/Frame 058363/0723 →
RELEASE OF SECURITY INTEREST Recorded Nov 30, 2021
From: CORTLAND CAPITAL MARKETS SERVICES LLC
To: PROVENANCE ASSET GROUP HOLDINGS LLC; PROVENANCE ASSET GROUP LLC
Reel/Frame 058983/0104 →
ASSIGNMENT AND ASSUMPTION AGREEMENT Recorded Feb 14, 2019
From: NOKIA USA INC.
To: NOKIA US HOLDINGS INC.
Reel/Frame 048370/0682 →
SECURITY INTEREST Recorded Sep 13, 2017
From: PROVENANCE ASSET GROUP HOLDINGS, LLC; PROVENANCE ASSET GROUP, LLC
To: CORTLAND CAPITAL MARKET SERVICES, LLC
Reel/Frame 043967/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 13, 2017
From: NOKIA TECHNOLOGIES OY; NOKIA SOLUTIONS AND NETWORKS BV; ALCATEL LUCENT SAS
To: PROVENANCE ASSET GROUP LLC
Reel/Frame 043877/0001 →
SECURITY INTEREST Recorded Sep 13, 2017
From: PROVENANCE ASSET GROUP HOLDINGS, LLC; PROVENANCE ASSET GROUP LLC
To: NOKIA USA INC.
Reel/Frame 043879/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 21, 2015
From: MARILLY, EMMANUEL; MIHOUB, ALAEDDINE; OUTTAGARTS, ABDELKADER
To: ALCATEL LUCENT
Reel/Frame 036389/0338 →
Priority Claims (1)
EP 12182267 · Aug 29, 2012 · regional
Continuity (1)
Related Publication 20150264306A1 · Sep 17, 2015