IP Library Granted Patent US 11,070,768
Granted Patent B1
US 11,070,768 · App. 17/075,408 · Granted Jul 20, 2021

Volume areas in a three-dimensional virtual conference space, and applications thereof

Inventors: Gerard Cornelis Krol (Leiden, NL); Erik Stuart Braund (Saugerties, NY)
Assignee: Katmai Tech Holdings LLC
H04N7/157G06F3/162G06F3/165G06T15/20
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,070,768
App. No.
17/075,408
Granted
Jul 20, 2021
Kind
B1
Abstract

Disclosed herein is a web-based videoconference system that allows for video avatars to navigate within the virtual environment. The system has a presented mode that allows for a presentation stream to be texture mapped to a presenter screen situated within the virtual environment. The relative left-right sound is adjusted to provide sense of an avatar's position in a virtual space. The sound is further adjusted based on the area where the avatar is located and where the virtual camera is located. Video stream quality is adjusted based on relative position in a virtual space. Three-dimensional modeling is available inside the virtual video conferencing environment.

Claims (48)

1. A computer-implemented method for providing audio for a virtual conference, comprising:

(a) from a perspective of a virtual camera of a first user, rendering for display to the first user at least a portion of a three-dimensional virtual space, the three-dimensional virtual space including an avatar representing a second user, the virtual camera at a first position in the three-dimensional virtual space and the avatar at a second position in the three-dimensional virtual space, wherein the three-dimensional virtual space is segmented into a plurality of areas;

(b) receiving an audio stream from a microphone of a device of the second user, the microphone positioned to capture speech of the second user;

(c) determining whether the virtual camera and the avatar are located in a same area in the plurality of areas;

(d) determining whether the avatar is in a podium area in the plurality of areas;

(e) when the virtual camera and the avatar are determined not to be located in the same area and the avatar is determined not to be in the podium area, attenuating the audio stream; and

(f) outputting the audio stream to be played to the first user.

2. The computer-implemented method of claim 1 , wherein the audio stream is a first audio stream, wherein the three-dimensional virtual space includes a second avatar representing a third user, wherein the determining (c) comprises determining that the virtual camera and the avatar are located in the same area, further comprising:

(g) receiving a second audio stream from a microphone of a device of the first user, the microphone positioned to capture speech of the first user;

(h) determining that the second avatar and the virtual camera are located in an area of the three-dimensional virtual space different from the same area where the virtual camera is located; and

(i) attenuating the first and second audio streams to prevent the first and second audio streams from being heard by the third user, enabling a private conversation between the first and second user.

3. The computer-implemented method of claim 1 , wherein respective areas in the plurality of areas have a wall transmission factor that specifies how much the audio stream is attenuated in (e).

4. The computer-implemented method of claim 1 , wherein respective areas in the plurality of areas have a distance transmission factor, further comprising:

(g) determining a distance in the three-dimensional virtual space between the virtual camera and the avatar;

(h) determining at least one area between the virtual camera and the avatar; and

(i) attenuating the audio stream based on the distance determined in (g) and the distance transmission factor corresponding to the at least one area determined in (h).

5. The computer-implemented method of claim 1 , wherein the plurality of areas is structured as a hierarchy.

6. The computer-implemented method of claim 5 , wherein the respective areas in the plurality of areas have a wall transmission factor, further comprising:

(g) traversing the hierarchy to determine a subset of areas from the plurality of areas between an area including the avatar and an area including the virtual camera; and

(h) attenuating the audio stream based on the respective wall transmission factor corresponding to the subset of areas determined in (g).

7. A non-transitory, tangible computer-readable device having instructions stored thereon that, when executed by at least one computing device, causes the at least one computing device to perform operations for providing audio for a virtual conference, comprising:

(a) from a perspective of a virtual camera of a first user, rendering for display to the first user at least a portion of a three-dimensional virtual space, the three-dimensional virtual space including an avatar representing a second user, the virtual camera at a first position in the three-dimensional virtual space and the avatar at a second position in the three-dimensional virtual space, wherein the three-dimensional virtual space is segmented into a plurality of areas;

(b) receiving an audio stream from a microphone of a device of the second user, the microphone positioned to capture speech of the second user;

(c) determining whether the virtual camera and the avatar are located in a same area in the plurality of areas;

(d) determining whether the avatar is in a podium area in the plurality of areas;

(e) when the virtual camera and the avatar are determined not to be located in the same area and the avatar is determined not to be in the podium area, attenuating the audio stream; and

(f) outputting the audio stream to be played to the first user.

8. The device of claim 7 , wherein the audio stream is a first audio stream, wherein the three-dimensional virtual space includes a second avatar representing a third user, wherein the determining (c) comprises determining that the virtual camera and the avatar are located in the same area, further comprising:

(g) receiving a second audio stream from a microphone of a device of the first user, the microphone positioned to capture speech of the first user;

(h) determining that the second avatar and the virtual camera are located in an area of the three-dimensional virtual space different from the same area where the virtual camera is located; and

(i) attenuating the first and second audio stream to prevent the first and second audio streams from being heard by the third user, enabling a private conversation between the first and second user.

9. The device of claim 7 , wherein respective areas in the plurality of areas have a wall transmission factor that specifies how much the audio stream is attenuated in (e).

10. The device of claim 7 , wherein respective areas in the plurality of areas have a distance transmission factor, further comprising:

(g) determining a distance in the three-dimensional virtual space between the virtual camera and the avatar;

(h) determining at least one area between the virtual camera and the avatar; and

(i) attenuating the audio stream based on the distance determined in (g) and the distance transmission factor corresponding to the at least one area determined in (h).

11. The device of claim 7 , wherein the plurality of areas is structured as a hierarchy.

12. The device of claim 11 , wherein the respective areas in the plurality of areas have a wall transmission factor, further comprising:

(g) traversing the hierarchy to determine a subset of areas from the plurality of areas between an area including the avatar and an area including the virtual camera; and

(h) attenuating the audio stream based on the respective wall transmission factor corresponding to the subset of areas determined in (g).

13. A system for providing audio for a virtual conference, comprising:

a processor coupled to memory;

a renderer implemented on the processor and configured to, from a perspective of a virtual camera of a first user, render for display to the first user at least a portion of a three-dimensional virtual space, the three-dimensional virtual space including an avatar representing a second user, the virtual camera at a first position in the three-dimensional virtual space and the avatar at a second position in the three-dimensional virtual space, wherein the three-dimensional virtual space is segmented into a plurality of areas structured as a hierarchy;

a network interface configured to receive an audio stream from a microphone of a device of the second user, the microphone positioned to capture speech of the second user; and

an audio processor configured to determine whether the virtual camera and the avatar are located in a same area in the plurality of areas and whether the avatar is in a podium area in the plurality of areas, and, when the virtual camera and the avatar are determined not to be located in the same area and the avatar is determined not to be in the podium area, attenuate the audio stream, and output the audio stream to be played to the first user.

14. The system of claim 13 , wherein respective areas in the plurality of areas have a wall transmission factor that specifies how much the audio stream is attenuated.

15. The system of claim 13 , wherein respective areas in the plurality of areas have a distance transmission factor, the audio processor is configured to: (i) determine a distance in the three-dimensional virtual space between the virtual camera and the avatar, (ii) determine at least one area between the virtual camera and the avatar, and (iii) attenuating the audio stream based on the determined distance and the distance transmission factor corresponding to the determined at least one area.

16. The system of claim 13 , wherein the respective areas in the plurality of areas have a wall transmission factor, wherein the audio processor is configured to traverse the hierarchy to determine a subset of areas from the plurality of areas between an area including the avatar and an area including the virtual camera, and attenuating the audio stream based on the respective wall transmission factor corresponding to the determined subset of areas determined.

Assignments (4)
CORRECTIVE ASSIGNMENT TO CORRECT THE ASSIGNEE SUITE NUMBER FROM "SUITE 410" PREVIOUSLY RECORDED ON REEL 054983 FRAME 0449. ASSIGNOR(S) HEREBY CONFIRMS THE CORRECT ASSIGNEE SUITE NUMBER IS --SUITE 310--. Recorded Jan 27, 2021
From: BRAUND INC.
To: KATMAI TECH HOLDINGS LLC
Reel/Frame 055132/0588 →
CORRECTIVE ASSIGNMENT TO CORRECT THE ASSIGNEE NAME PREVIOUSLY RECORDED AT REEL: 054132 FRAME: 0743. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Jan 26, 2021
From: KROL, GERARD CORNELIS; BRAUND, ERIK STUART
To: BRAUND INC.
Reel/Frame 055122/0889 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 21, 2021
From: BRAUND INC.
To: KATMAI TECH HOLDINGS LLC
Reel/Frame 054983/0449 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 21, 2020
From: KROL, GERARD CORNELIS; BRAUND, ERIK STUART
To: BRAUND, INC.
Reel/Frame 054132/0743 →
Cited By (6)
US 12,192,679 US 12,229,912 US 12,293,477 US 12,340,585 US 12,413,687 US 12,489,652