IP Library Granted Patent US 12,474,815
Granted Patent B2
US 12,474,815 · App. 18/741,135 · Granted Nov 18, 2025

Graphical user interface configuration for display at an output interface during a video conference

Inventor: Cary Arnold Bran (Vashon, WA)
Assignee: Zoom Communications, Inc.
G06F3/04815G06F3/04845
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,474,815
App. No.
18/741,135
Granted
Nov 18, 2025
Kind
B2
Abstract

A graphical user interface (GUI) may be configured for display at an output interface during a video conference. The GUI may comprise visual elements associated with participant devices of the video conference. For example, the visual elements may include video feeds and/or images associated with the participant devices. During the video conference, a first visual element may be moved to a location in the GUI based on a characteristic associated with the first visual element. The characteristic and the location may be based on an input. In some implementations, the visual elements may be arranged in a two-dimensional visual layout. In some implementations, the visual elements may be arranged in a three-dimensional visual layout. The visual elements may be moved, for example, based on a communication sent during the video conference, an arrival of a participant to the video conference, and/or a communication modality used during the video conference.

Claims (57)

1 . A method, comprising:

receiving an input from a first participant device to select a template that defines characteristics associated with visual elements and defines locations in a graphical user interface (GUI);

detecting, during a video conference, a first characteristic of the characteristics associated with a second participant device and a first visual element of the visual elements, wherein the first characteristic is an arrival, to the video conference, of a participant associated with the second participant device;

moving, during the video conference, the first visual element to a first location in the GUI based on the first characteristic, wherein the first location overlaps with a second location in the GUI associated with a second visual element of the visual elements;

adjusting an audio level associated with the second visual element based on the first location overlapping with the second location;

causing, based on detecting the first characteristic, a third visual element to appear at a third location of the GUI, wherein the third visual element is a participant roster; and

fading, based on not detecting an arrival of a participant to the video conference for a period of time, the third visual element.

2 . The method of claim 1 , wherein the visual elements include shared content.

3 . The method of claim 1 , further comprising:

adjusting an audio level associated with the first visual element based on the first location overlapping with the second location.

4 . The method of claim 1 , wherein the visual elements are arranged in a three-dimensional visual layout and the adjusting the audio level comprises adjusting the audio level in a physical meeting room based on spatial mapping of the first visual element and the second visual element.

5 . The method of claim 1 , further comprising:

displaying the first visual element in a height dimension and a width dimension while moving the first visual element in a depth dimension.

6 . The method of claim 1 , further comprising:

determining a first position information and a second position information for displaying the first visual element and the second visual element of the visual elements in a three-dimensional visual layout, respectively; and

calculating a change to the first position information to move the first visual element in front of or behind the second visual element in the three-dimensional visual layout.

7 . The method of claim 1 , wherein the visual elements are arranged in a three-dimensional visual layout.

8 . The method of claim 1 , further comprising:

arranging the first visual element at an angle relative to the second visual element of the visual elements in a three-dimensional visual layout.

9 . The method of claim 1 , further comprising:

moving the second visual element of the visual elements to a third location based on detecting a second characteristic associated with the second visual element; and

ordering the second visual element relative to the first visual element at the first location.

10 . An apparatus, comprising:

a memory; and

a processor configured to execute instructions stored in the memory to:

receive an input from a first participant device to select a template that defines characteristics associated with visual elements and defines locations in a graphical user interface (GUI);

detect, during a video conference, a first characteristic of the characteristics associated with a second participant device and a first visual element of the visual elements, wherein the first characteristic is an arrival, to the video conference, of a participant associated with the second participant device;

move, during the video conference, the first visual element to a first location in the GUI based on the first characteristic, wherein the first location overlaps with a second location in the GUI associated with a second visual element of the visual elements;

adjust an audio level associated with the second visual element based on the first location overlapping with the second location;

cause, based on the detection of the first characteristic, a third visual element to appear at a third location of the GUI, wherein the third visual element is a participant roster; and

fade, based on a non-detection of an arrival of a participant to the video conference for a period of time, the third visual element.

11 . The apparatus of claim 10 , wherein the visual elements include a video feed.

12 . The apparatus of claim 10 , wherein the processor is further configured to execute instructions stored in the memory to:

display the first visual element in a height dimension and a width dimension while moving the first visual element in a depth dimension.

13 . The apparatus of claim 10 , wherein the processor is further configured to execute instructions stored in the memory to:

determine a first position information and a second position information to display the first visual element and the second visual element of the visual elements in a three-dimensional visual layout, respectively; and

calculate a change to the first position information to move the first visual element in front of or behind the second visual element in the three-dimensional visual layout.

14 . The apparatus of claim 10 , wherein the processor is further configured to execute instructions stored in the memory to:

at least one of resize or fade the first visual element when moving the first visual element.

15 . A non-transitory computer-readable medium storing instructions operable to cause one or more processors to perform operations comprising:

receiving an input from a first participant device to select a template that defines characteristics associated with visual elements and defines locations in a graphical user interface (GUI);

detecting, during a video conference, a first characteristic of the characteristics associated with a second participant device and a first visual element of the visual elements, wherein the first characteristic is an arrival, to the video conference, of a participant associated with the second participant device;

moving, during the video conference, the first visual element to a first location in the GUI based on the first characteristic, wherein the first location overlaps with a second location in the GUI associated with a second visual element of the visual elements;

adjusting an audio level associated with the second visual element based on the first location overlapping with the second location;

causing, based on detecting the first characteristic, a third visual element to appear at a third location of the GUI, wherein the third visual element is a participant roster; and

fading, based on not detecting an arrival of a participant to the video conference for a period of time, the third visual element.

16 . The non-transitory computer-readable medium storing instructions of claim 15 , the operations further comprising:

determining the first characteristic and the first location based on the input, wherein the input indicates a first selection of the first characteristic and a second selection of the first location; and

arranging the first visual element in a two-dimensional visual layout.

17 . The non-transitory computer-readable medium storing instructions of claim 15 , the operations further comprising:

selecting a second template that indicates a second characteristic and the second location, wherein the input indicates a selection of the second template among multiple templates.

18 . The non-transitory computer-readable medium storing instructions of claim 15 , wherein the visual elements are arranged in a three-dimensional visual layout and the adjusting the audio level comprises adjusting the audio level in a physical meeting room based on spatial mapping of the first visual element and the second visual element.

19 . The non-transitory computer-readable medium storing instructions of claim 15 , the operations further comprising:

determining a first position information and a second position information for displaying the first visual element and the second visual element of the visual elements in a three-dimensional visual layout, respectively; and

calculating a change to the first position information to move the first visual element in front of or behind the second visual element in the three-dimensional visual layout.

20 . The non-transitory computer-readable medium storing instructions of claim 15 , the operations further comprising:

adjusting an audio level associated with the first visual element based on the first location overlapping with the second location.

Assignments (2)
CHANGE OF NAME Recorded Jan 7, 2025
From: ZOOM VIDEO COMMUNICATIONS, INC.
To: ZOOM COMMUNICATIONS, INC.
Reel/Frame 069839/0593 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 12, 2024
From: BRAN, CARY ARNOLD
To: ZOOM VIDEO COMMUNICATIONS, INC.
Reel/Frame 067709/0244 →
Continuity (2)
Continuation 17728615 · Apr 25, 2022
Related Publication 20240329798A1 · Oct 3, 2024
References Cited (21)
US 8537196B2 · Hegde et al. · 2013 [cited by applicant]
US 8797377B2 · Mauchly et al. · 2014 [cited by applicant]
US 9813673B2 · Smits · 2017 [cited by applicant]
US 9924136B1 · Faulkner · 2018 [cited by examiner]
US 10061467B2 · Brunsch et al. · 2018 [cited by applicant]
US 10509964B2 · Astavans et al. · 2019 [cited by applicant]
US 10750124B2 · Rosenberg · 2020 [cited by applicant]
US 10880582B2 · Goldman et al. · 2020 [cited by applicant]
US 11394925B1 · Faulkner et al. · 2022 [cited by applicant]
US 20020093531A1 · Barile · 2002 [cited by applicant]
US 20050099492A1 · Orr · 2005 [cited by applicant]
US 20070171275A1 · Kenoyer · 2007 [cited by applicant]
US 20100085416A1 · Hegde et al. · 2010 [cited by applicant]
US 20120274736A1 · Robinson · 2012 [cited by examiner]
US 20160308920A1 · Brunsch et al. · 2016 [cited by applicant]
US 20170351476A1 · Yoakum · 2017 [cited by applicant]
US 20200371677A1 · Faulkner · 2020 [cited by examiner]
US 20220103963A1 · Satongar · 2022 [cited by examiner]
US 20220150288A1 · Tokuchi · 2022 [cited by applicant]
US 20230044865A1 · Pitts et al. · 2023 [cited by applicant]
EP 2614637B1 · 2018 [cited by applicant]