Conference system for use of multiple devices
A conference system is described that associates a first device and a second device to the same user, compares a first input from the first device and a second input from the second device, and modifies a setting of a conference session. The first input and the second input may be a video input or an audio input. The modification may include, for example, noise removal, determination of the user's AV feed device, or removing a background image.
1 . A computer-implemented method for a conference system, the computer-implemented method comprising:
assigning, with a same user, a first device and a second device participating in a conference session, wherein assigning comprises detecting that the same user logs into the conference session using both the first device and the second device;
comparing a first audio input from the first device and a second audio input from the second device, wherein the conference system receives the first audio input and the second audio input during the conference session; and
modifying a setting of the conference session based on a result of the comparison of the first audio input and the second audio input, wherein modifying comprises designating either the first audio input or the second audio input as an audio feed for the same user for the conference session, and when the audio feed is the first audio input, removing from the audio feed a sound that arrives later at the first device than at the second device.
2 . The computer-implemented method of claim 1 , wherein:
comparing the first audio input from the first device and the second audio input from the second device further comprises comparing a first time when a sound is detected on the first device with a second time when the sound is detected on the second device; and
modifying the setting of the conference session based on the result of the comparison of the first audio input and the second audio input further comprises removing from the audio feed a sound that arrives later at the second device than at the first device, when the audio feed is the second audio input.
3 . The computer-implemented method of claim 1 , wherein:
comparing the first audio input from the first device and the second audio input from the second device further comprises comparing a first time when a sound is detected on the first device with a second time when the sound is detected on the second device; and
modifying the setting of the conference session based on the result of the comparison of the first audio input and the second audio input further comprises setting a video input of the first device as the user's video input where the first time is earlier than the second time.
4 . The computer-implemented method of claim 1 , wherein:
comparing the first audio input from the first device and the second audio input from the second device further comprises comparing a first time when a sound is detected on the first device with a second time when the sound is detected on the second device to determine a location of the sound's source; and
modifying the setting of the conference session based on the result of the comparison of the first audio input and the second audio input further comprises removing the sound from one of the first device and the second device based on the location of the sound's source.
5 . The computer-implemented method of claim 1 , wherein:
the first input further comprises a first video input from the first device and the second input further comprises a second video input from the second device.
6 . The computer-implemented method of claim 5 , wherein:
comparing the first input from the first device and the second input from the second device comprises comparing a first image captured from the first video input with a second image captured from the second video input to determine a depth of an object in the first image and the second image; and
modifying the setting of the conference session based on the result of the comparison of the first input and the second input comprises applying a background image processing to the object to at least one of the first video input and the second video input based on the depth of the object.
7 . The computer-implemented method of claim 5 , wherein:
comparing the first input from the first device and the second input from the second device further comprises comparing a first direction of the user's line of sight in the first video input and a second direction of the user's line of sight in the second video input; and
modifying the setting of the conference session based on the result of the comparison of the first input and the second input further comprises setting a video input of the first device as the user's video input when the user is determined to be more directly facing the first device based on the first direction and the second direction.
8 . A conference system, comprising:
a memory configured to store operations; and
one or more processors configured to perform the operations, the operations comprising:
assigning, with a same user, a first device and a second device participating in a conference session, wherein assigning comprises detecting that the same user logs into the conference session using both the first device and the second device;
comparing a first audio input from the first device and a second audio input from the second device, wherein the conference system receives the first audio input and the second audio input during the conference session; and
modifying a setting of the conference session based on a result of the comparison of the first audio input and the second audio input, wherein modifying comprises designating either the first audio input or the second audio input as an audio feed for the same user for the conference session, and when the audio feed is the first audio input, removing from the audio feed a sound that arrives later at the first device than at the second device.
9 . The system of claim 8 , wherein:
comparing the first audio input from the first device and the second audio input from the second device further comprises comparing a first time when a sound is detected on the first device with a second time when the sound is detected on the second device; and
modifying the setting of the conference session based on the result of the comparison of the first audio input and the second audio input further comprises removing from the audio feed a sound that arrives later at the second device than at the first device, when the audio feed is the second audio input.
10 . The system of claim 8 , wherein:
comparing the first audio input from the first device and the second audio input from the second device further comprises comparing a first time when a sound is detected on the first device with a second time when the sound is detected on the second device; and
modifying the setting of the conference session based on the result of the comparison of the first audio input and the second audio input further comprises setting a video input of the first device as the user's video input where the first time is earlier than the second time.
11 . The system of claim 8 , wherein:
comparing the first audio input from the first device and the second audio input from the second device further comprises comparing a first time when a sound is detected on the first device with a second time when the sound is detected on the second device to determine a location of the sound's source; and
modifying the setting of the conference session based on the result of the comparison of the first audio input and the second audio input further comprises removing the sound from one of the first device and the second device based on the location of the sound's source.
12 . The system of claim 8 , wherein:
the first input further comprises a first video input from the first device and the second input further comprises a second video input from the second device.
13 . The system of claim 12 wherein:
comparing the first input from the first device and the second input from the second device comprises comparing a first image captured from the first video input with a second image captured from the second video input to determine a depth of an object in the first image and the second image; and
modifying the setting of the conference session based on the result of the comparison of the first input and the second input comprises applying a background image processing to the object to at least one of the first video input and the second video input based on the depth of the object.
14 . The system of claim 12 , wherein:
comparing the first input from the first device and the second input from the second device further comprises comparing a first direction of the user's line of sight in the first video input and a second direction of the user's line of sight in the second video input; and
modifying the setting of the conference session based on the result of the comparison of the first input and the second input further comprises setting a video input of the first device as the user's video input when the user is determined to be more directly facing the first device based on the first direction and the second direction.
15 . A computer readable storage device having instructions stored thereon that, when executed by one or more processing devices, cause the one or more processing devices to perform operations comprising:
assigning, with a same user, a first device and a second device participating in a conference session, wherein assigning comprises detecting that the same user logs into the conference session using both the first device and the second device;
comparing a first audio input from the first device and a second audio input from the second device, wherein a conference system receives the first audio input and the second audio input during the conference session; and
modifying a setting of the conference session based on a result of the comparison of the first audio input and the second audio input, wherein modifying comprises designating either the first audio input or the second audio input as an audio feed for the same user for the conference session, and when the audio feed is the first audio input, removing from the audio feed a sound that arrives later at the first device than at the second device.
16 . The computer readable storage device of claim 15 , wherein:
comparing the first audio input from the first device and the second audio input from the second device further comprises comparing a first time when a sound is detected on the first device with a second time when the sound is detected on the second device; and
modifying the setting of the conference session based on the result of the comparison of the first audio input and the second audio input further comprises removing from the audio feed a sound that arrives later at the second device than at the first device, when the audio feed is the second audio input.
17 . The computer readable storage device of claim 15 , wherein:
comparing the first audio input from the first device and the second audio input from the second device further comprises comparing a first time when a sound is detected on the first device with a second time when the sound is detected on the second device; and
modifying the setting of the conference session based on the result of the comparison of the first audio input and the second audio input further comprises setting a video input of the first device as the user's video input where the first time is earlier than the second time.
18 . The computer readable storage device of claim 15 , wherein:
comparing the first audio input from the first device and the second audio input from the second device further comprises comparing a first time when a sound is detected on the first device with a second time when the sound is detected on the second device to determine a location of the sound's source; and
modifying the setting of the conference session based on the result of the comparison of the first audio input and the second audio input further comprises removing the sound from one of the first device and the second device based on the location of the sound's source.
19 . The computer readable storage device of claim 15 , wherein:
the first input further comprises a first video input from the first device and the second input further comprises a second video input from the second device;
comparing the first input from the first device and the second input from the second device comprises comparing a first image captured from the first video input with a second image captured from the second video input to determine a depth of an object in the first image and the second image; and
modifying the setting of the conference session based on the result of the comparison of the first input and the second input comprises applying a background image processing to the object to at least one of the first video input and the second video input based on the depth of the object.
20 . The computer readable storage device of claim 15 , wherein:
the first input further comprises a first video input from the first device and the second input further comprises a second video input from the second device;
comparing the first input from the first device and the second input from the second device further comprises comparing a first direction of the user's line of sight in the first video input and a second direction of the user's line of sight in the second video input; and
modifying the setting of the conference session based on the result of the comparison of the first input and the second input further comprises setting a video input of the first device as the user's video input when the user is determined to be more directly facing the first device based on the first direction and the second direction.