IP Library › Granted Patent US 12,317,002
Granted Patent B2
US 12,317,002 · App. 18/047,420 · Granted May 27, 2025

Selecting avatar for videoconference

Inventors: Yinda Zhang (Daly City, CA); Ruofei Du (San Francisco, CA)
Assignee: Google LLC
H04N7/157G06F3/167G06T13/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,317,002
App. No.
18/047,420
Granted
May 27, 2025
Kind
B2
Abstract

A method can include selecting, from at least a first avatar and a second avatar based on at least one attribute of a calendar event associated with a user, a session avatar, the first avatar being based on a first set of images of a user wearing a first outfit and the second avatar being based on a second set of images of the user wearing a second outfit, and presenting the session avatar during a videoconference, the presentation of the session avatar changing based on audio input received from the user during the videoconference.

Claims (42)

1. A method comprising:

storing an association between features for a plurality of avatars and probabilities for the plurality of avatars, the plurality of avatars including a first avatar that was generated based on a first set of images of a user wearing a first outfit and a second avatar that was generated based on a second set of images of the user wearing a second outfit, the features being based on at least one of the first set of images and the second set of images;

selecting the first avatar as a session avatar based on video input and the association between the features and the probabilities, the second avatar being an unselected avatar; and

presenting the session avatar during a videoconference.

2. The method of claim 1 , wherein the selection of the session avatar is further based on at least one attribute of a calendar event, a description of the first avatar received from the user, and a description of the second avatar received from the user.

3. The method of claim 2 , wherein:

the description of the first avatar includes a contact identifier; and

the at least one attribute of the calendar event includes the contact identifier.

4. The method of claim 1 , wherein the selection is further based on a time of day of a calendar event.

5. The method of claim 1 , wherein the selection of the session avatar is further based on at least one attribute of a calendar event associated with the videoconference.

6. The method of claim 1 , wherein the selection of the session avatar is further based on audio input received from the user.

7. The method of claim 1 , wherein the selection of the session avatar is further based on previous selections, by the user, of either the first avatar or the second avatar.

8. The method of claim 1 , wherein the selection of the session avatar is further based on video input of the user received during previous videoconferences, the video input indicating a type of outfit worn by the user during the previous videoconferences.

9. The method of claim 1 , wherein the video input indicates a type of outfit worn by the user during the videoconference.

10. The method of claim 1 , wherein the selection of the session avatar includes:

presenting the session avatar to the user; and

receiving an indication of approval of the session avatar from the user.

11. The method of claim 1 , further comprising:

capturing a third set of images of the user during the videoconference; and

storing a third avatar based on the third set of images of the user.

12. The method of claim 1 , wherein the user was speaking predetermined words from a script as part of generating the first avatar.

13. The method of claim 1 , wherein selecting the first avatar as the session avatar is further based on a tone of voice of the user.

14. The method of claim 1 , further comprising:

receiving audio input during presentation of the session avatar during the videoconference; and

changing the session avatar to the second avatar based on features in the audio input.

15. The method of claim 1 , wherein the features include facial landmarks.

16. The method of claim 1 , wherein the association between the features and the probabilities are stored in a model.

17. A non-transitory computer-readable storage medium comprising instructions stored thereon that, when executed by at least one processor, are configured to cause a computing system to:

store an association between features for a plurality of avatars and probabilities for the plurality of avatars, the plurality of avatars including a first avatar that was generated based on a first set of images of a user wearing a first outfit and a second avatar that was generated based on a second set of images of the user wearing a second outfit, the features being based on at least one of the first set of images and the second set of images;

select the first avatar as a session avatar based on video input and the association between the features and the probabilities, the second avatar being unselected; and

present the session avatar during a videoconference.

18. The non-transitory computer-readable storage medium of claim 17 , wherein the selection of the session avatar is further based on at least one attribute of a calendar event, a description of the first avatar received from the user, and a description of the second avatar received from the user.

19. The non-transitory computer-readable storage medium of claim 17 , wherein the selection of the session avatar is further based on an attendee of a calendar event other than the user.

20. A computing system comprising:

at least one processor; and

a non-transitory computer-readable storage medium comprising instructions stored thereon that, when executed by the at least one processor, are configured to cause the computing system to:

store an association between features for a plurality of avatars and probabilities for the plurality of avatars, the plurality of avatars including a first avatar that was generated based on a first set of images of a user wearing a first outfit and a second avatar that was generated based on a second set of images of the user wearing a second outfit, the features being based on at least one of the first set of images and the second set of images;

select the first avatar as based on video input and the association between the features and the probabilities, the second avatar being an unselected avatar; and

present the session avatar during a videoconference.

21. The computing system of claim 20 , wherein the selection of the first avatar is further based on an attendee of a calendar event other than the user.

22. The computing system of claim 20 , wherein the instructions are further configured to cause the computing system to:

change the session avatar to the second avatar based on features in audio input received during presentation of the session avatar during the videoconference.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 20, 2022
From: ZHANG, YINDA; DU, RUOFEI
To: GOOGLE LLC
Reel/Frame 061476/0446 →
Continuity (1)
Related Publication 20240129437A1 · Apr 18, 2024
References Cited (15)
US 20090158150A1 · Lyle · 2009 [cited by examiner]
US 20090300525A1 · Jolliff · 2009 [cited by examiner]
US 20100201693A1 · Caplette · 2010 [cited by examiner]
US 20110014932A1 · Estevez · 2011 [cited by examiner]
US 20120089908A1 · Miyaki · 2012 [cited by examiner]
US 20140036027A1 · Liu · 2014 [cited by examiner]
US 20150302662A1 · Miller · 2015 [cited by applicant]
US 20180373547A1 · Dawes · 2018 [cited by examiner]
US 20200202603A1 · Choi et al. · 2020 [cited by applicant]
US 20210090314A1 · Hussen Abdelaziz et al. · 2021 [cited by applicant]
US 20210104087A1 · Smith et al. · 2021 [cited by applicant]
US 20210358190A1 · Choi · 2021 [cited by examiner]
US 20210405831A1 · Mourkogiannis et al. · 2021 [cited by applicant]
WO WO2019008320A1 · 2019 [cited by examiner]
Nguyen, et al., “Automatic Generation of a 3D Sign Language Avatar on AR Glasses Given 2D Videos of Human Signers”, Proceedings of the 18th Biennial Machine Translation Summit, Virtual USA, 1st International Workshop on… [cited by applicant]