IP Library Granted Patent US 11,743,430
Granted Patent B2
US 11,743,430 · App. 17/313,338 · Granted Aug 29, 2023

Providing awareness of who can hear audio in a virtual conference, and applications thereof

Inventors: Gerard Cornelis Krol (Leiden, NL); Erik Stuart Braund (Saugerties, NY)
Assignee: Katmai Tech Inc.
H04N7/157G06T15/04G06T15/20G10L25/51
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,743,430
App. No.
17/313,338
Granted
Aug 29, 2023
Kind
B2
Abstract

Disclosed herein is a computer-implemented method, system, device, and computer program product for providing awareness of who are able to hear audio in a virtual conference. For each of the users in the virtual conference, a device of the speaking user or a server determines whether a respective user is able to hear the speaking user based on whether a respective sound volume at which the respective user is able hear the speaking user exceeds a threshold amount. The speaking user's device or the server outputs a notification indicating the users that are able to hear the speaking user.

Claims (78)

1. A computer-implemented method for providing awareness of who can hear audio in a virtual conference, comprising:

(a) from a perspective of a virtual camera of a speaking user, rendering for display to the speaking user a three-dimensional virtual space including a plurality of avatars, wherein each user of a plurality of users is represented by a respective avatar of the plurality of avatars;

for each user of the plurality of users:

(b) determining a sound volume based on a position in the three-dimensional virtual space of the respective avatar relative to a position of the virtual camera in the three-dimensional virtual space; and

(c) determining whether a user is able to audibly hear the speaking user based on whether the sound volume exceeds a threshold amount, wherein the sound volume can be higher or lower than the threshold amount;

(d) generating a list of users from the plurality of users such that the list of users consists of users that are able to audibly hear the speaking user based on whether the sound volume for each respective user in the list of users exceeds the threshold amount; and

(e) outputting a single visual notification of the list of users for display to the speaking user, wherein the single visual notification of the list of users excludes users from the plurality of users that are unable to audibly hear the speaking user.

2. The computer-implemented method of claim 1 , wherein the visual notification includes a total count of the list of users determined in (c) that are able to audibly hear the speaking user.

3. The computer-implemented method of claim 1 , further comprising:

(f) receiving position and direction information of at least one user of the plurality of users in the three-dimensional virtual space;

(g) receiving a video stream captured from a camera of a device of the at least one user, the camera positioned to capture photographic images of the at least one user;

(h) texture mapping the video stream onto a three-dimensional model of an avatar of the at least one user; and

(i) rendering, from the perspective of the virtual camera of the speaking user, for display to the speaking user, the three-dimensional virtual space including the texture-mapped three-dimensional model of the avatar located at a position and oriented at a direction based on the position and direction information of the at least one user.

4. The computer-implemented method of claim 1 , wherein an application downloaded on a web-browser renders the three-dimensional virtual space in (a).

5. The computer-implemented method of claim 1 , wherein the three-dimensional virtual space is segmented into a plurality of areas and each area in the plurality of areas comprises a wall transmission factor and a distance transmission factor.

6. The computer-implemented method of claim 5 , wherein the determining of the sound volume in (b) is based on one or more of:

i) a distance between the virtual camera and the respective avatar, and the wall transmission factors of all areas between the virtual camera and the respective avatar; or

ii) the distance transmission factors of all areas between the virtual camera and the respective avatar.

7. The computer-implemented method of claim 1 , further comprising:

(f) determining whether a new user is able to audibly hear the speaking user based on a new sound volume for the new user exceeding the threshold amount; and

(g) adding information about the new user in the visual notification.

8. The computer-implemented method of claim 1 , further comprising:

(f) determining that at least one sound volume associated with at least one user in the plurality of users has dropped lower than the threshold amount; and

(g) updating information about the at least one user from the visual notification.

9. A non-transitory, tangible computer-readable device having instructions stored thereon that, when executed by at least one computing device, causes the at least one computing device to perform operations for providing awareness of who can hear audio in a virtual conference, comprising:

(a) from a perspective of a virtual camera of a speaking user, rendering for display to the speaking user a three-dimensional virtual space, including a plurality of avatars, wherein each user of a plurality of users is represented by a respective avatar of the plurality of avatars;

for each user of the plurality of users:

(b) determining a sound volume based on a position in the three-dimensional virtual space of the respective avatar relative to a position of the virtual camera in the three-dimensional virtual space; and

(c) determining whether a user is able to audibly hear the speaking user based on whether the sound volume exceeds a threshold amount, wherein the sound volume can be higher or lower than the threshold amount;

(d) generating a list of users from the plurality of users such that the list of users consists of users that are able to audibly hear the speaking user based on whether the sound volume for each respective user in the list of users exceeds the threshold amount; and

(e) outputting a single visual notification of the list of users for display to the speaking user, wherein the single visual notification of the list of users excludes users from the plurality of users that are unable to audibly hear the speaking user.

10. The non-transitory, tangible computer-readable device of claim 9 , wherein the visual notification includes a total count of the list of users determined in (c) that are able to audibly hear the speaking user.

11. The non-transitory, tangible computer-readable device of claim 9 , wherein the operations further comprise:

(f) receiving position and direction information of at least one user of the plurality of users in the three-dimensional virtual space;

(g) receiving a video stream captured from a camera of a device of the at least one user, the camera positioned to capture photographic images of the at least one user;

(h) texture mapping the video stream onto a three-dimensional model of an avatar of the at least one user; and

(i) rendering, from the perspective of the virtual camera of the speaking user, for display to the speaking user, the three-dimensional virtual space including the texture-mapped three-dimensional model of the avatar located at a position and oriented at a direction based on the position and direction information of the at least one user.

12. The non-transitory, tangible computer-readable device of claim 9 , wherein an application downloaded on a web-browser renders the three-dimensional virtual space in (a).

13. The non-transitory, tangible computer-readable device of claim 9 , wherein the three-dimensional virtual space is segmented into a plurality of areas and each area in the plurality of areas comprises a wall transmission factor and a distance transmission factor.

14. The non-transitory, tangible computer-readable device of claim 13 , wherein the determining of the sound volume in (b) is based on one or more of:

i) a distance between the virtual camera and the respective avatar, and the wall transmission factors of all areas between the virtual camera and the respective avatar; or

ii) the distance transmission factors of all areas between the virtual camera and the respective avatar.

15. The non-transitory, tangible computer-readable device of claim 9 , wherein the operations further comprise:

(f) determining whether a new user is able to audibly hear the speaking user based on a new sound volume for the new user exceeding the threshold amount; and

(g) adding information about the new user in the visual notification.

16. The non-transitory, tangible computer-readable device of claim 9 , wherein the operations further comprise:

(f) determining that at least one sound volume associated with at least one user in the plurality of users has dropped lower than the threshold amount; and

(g) updating information about the at least one user from the visual notification.

17. A system for providing awareness of who can hear audio in a virtual conference, the system comprising:

a memory; and

a processor coupled to the memory, the processor configured to:

(a) from a perspective of a virtual camera of a speaking user, render for display to the speaking user a three-dimensional virtual space, including a plurality of avatars, wherein each user of a plurality of users is represented by a respective avatar of the plurality of avatars;

for each user of the plurality of users:

(b) determine a sound volume based on a position in the three-dimensional virtual space of the respective avatar relative to a position of the virtual camera in the three-dimensional virtual space; and

(c) determine whether a user is able to audibly hear the speaking user based on whether the respective sound volume exceeds a threshold amount, wherein the sound volume can be higher or lower than the threshold amount;

(d) generate a list of users from the plurality of users such that the list of users consists of users that are able to audibly hear the speaking user based on whether the sound volume for each respective user in the list of users exceeds the threshold amount; and

(e) output a single visual notification of the list of users for display to the speaking user, wherein the single visual notification of the list of users excludes users from the plurality of users that are unable to audibly hear the speaking user.

18. The system of claim 17 , wherein the processor is further configured to:

(f) receive position and direction information of at least one user of the plurality of users in the three-dimensional virtual space;

(g) receive a video stream captured from a camera of a device of the at least one user, the camera positioned to capture photographic images of the at least one user;

(h) texture map the video stream onto a three-dimensional model of an avatar of the at least one user; and

(i) render, from the perspective of the virtual camera of the speaking user, for display to the speaking user, the three-dimensional virtual space including the texture-mapped three-dimensional model of the avatar located at a position and oriented at a direction based on the position and direction information of the at least one user.

19. The system of claim 17 , wherein an application downloaded on a web-browser renders the three-dimensional virtual space in (a).

20. The system of claim 17 , wherein the three-dimensional virtual space is segmented into a plurality of areas and each area in the plurality of areas has a wall transmission factor and a distance transmission factor, and

wherein the determining of the sound volume in (b) is based on one or more of:

i) a distance between the virtual camera and the respective avatar, and the wall transmission factors of all areas between the virtual camera and the respective avatar; or

ii) the distance transmission factors of all areas between the virtual camera and the respective avatar.

21. A computer-implemented method for providing awareness of who can hear audio in a virtual conference, comprising:

(a) from a perspective of a virtual camera of a speaking user, rendering for display to the speaking user a three-dimensional virtual space, including a plurality of avatars, wherein each user of a plurality of users is represented by a respective avatar of the plurality of avatars;

for each user of the plurality of users:

(b) determining whether a user is able to audibly hear the speaking user based on information associated with the user; and

(c) generating a list of users from the plurality of users such that the list of users consists of users that are able to audibly hear the speaking user based on the information associated with the user for each respective user in the list of users; and

(d) outputting a single visual notification of the list of users for display to the speaking user, wherein the single visual notification of the list of users excludes users from the plurality of users that are unable to audibly hear the speaking user.

22. The computer-implemented method of claim 21 , wherein determining that the user is able to audibly hear the speaking user in (b) based on the information associated with the user indicating that:

the user is in a predetermined group of users able to audibly hear the speaking user;

the user is of a predetermined user role; or

the user has predetermined security credentials.

23. The computer-implemented method of claim 21 , wherein determining that the user is not able to audibly hear the speaking user in (b) based on the information associated with the user indicating that the user is actively blocked by the speaking user from receiving the audio.

Assignments (3)
MERGER Recorded May 5, 2022
From: KATMAI TECH HOLDINGS LLC
To: KATMAI TECH INC.
Reel/Frame 059829/0335 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 3, 2021
From: KROL, GERARD CORNELIS; BRAUND, ERIK STUART
To: BRAUND, INC.
Reel/Frame 057065/0290 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 3, 2021
From: BRAUND INC.
To: KATMAI TECH HOLDINGS LLC
Reel/Frame 057065/0507 →
Continuity (1)
Related Publication 20220360742A1 · Nov 10, 2022
Cited By (1)
US 12,603,971