IP Library Granted Patent US 12,627,770
Granted Patent B2
US 12,627,770 · App. 18/033,855 · Granted May 12, 2026

Providing a 3D representation of a transmitting participant in a virtual meeting

Inventors: Ali El Essaili (Aachen, DE); Natalya Tyudina (Aachen, DE); Joerg Christian Ewert (Aachen, DE); Ola Melander (Wuerselen, DE)
Assignee: Telefonaktiebolaget LM Ericsson (publ)
H04N7/157G06V40/10G06V2201/12
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,627,770
App. No.
18/033,855
Granted
May 12, 2026
Kind
B2
Abstract

A method for providing a three-dimensional (3D) representation of a transmitting participant in a virtual meeting is provided. The method is performed in a representation provider and comprises obtaining a non-realtime 3D model of at least part of a person, obtaining partial realtime 3D data of the transmitting participant of the virtual meeting, and combining the non-realtime 3D model with the partial realtime 3D data, resulting in a combined 3D representation of the transmitting participant.

Claims (39)

1 . A method for providing a three-dimensional, 3D, representation of a transmitting participant in a virtual meeting, the method being performed in a representation provider, the method comprising:

obtaining a live camera stream representing parts of the transmitting participant of the virtual meeting;

based on a characteristic of the transmitting participant in the obtained live camera stream, selecting a non-realtime 3D model from a plurality of non-realtime 3D models that are general models that can be used for several different transmitting participants;

identifying at least one body feature in the non-realtime 3D model;

identifying at least one body feature in the live camera stream; and

generating a combined 3D representation of the transmitting participant based on both the non-realtime 3D model and the live camera stream, wherein generating the combined 3D representation further comprises, when the live camera stream is temporarily unavailable, combining the non-realtime 3D model with the most recently received live camera stream, and generating an appearance of the transmitting participant representing unavailability.

2 . The method according to claim 1 , further comprising:

generating the non-realtime 3D model based on camera data.

3 . The method according to claim 1 , wherein generating the combined 3D representation comprises, for body features forming part of both the non-realtime 3D model and the live camera stream, assigning higher weights to data included in the live camera stream than data included in the non-realtime 3D model.

4 . The method according to claim 1 , further comprising:

transmitting the combined 3D representation to a user device of the transmitting participant.

5 . The method according to claim 4 , further comprising:

receiving a performance indication from the user device of the transmitting participant.

6 . The method according to claim 1 , wherein the representation provider forms part of a user device of a receiving participant.

7 . The method according to claim 6 , wherein the user device of the receiving participant is an extended reality, XR, device.

8 . The method according to claim 1 , wherein the representation provider forms part of a server.

9 . The method according to claim 8 , further comprising:

transmitting the combined 3D representation to a user device of a receiving participant.

10 . A representation provider for providing a three-dimensional, 3D, representation of a transmitting participant in a virtual meeting, the representation provider comprising:

a processor; and

a memory storing instructions that, when executed by the processor, cause the representation provider to:

obtain a live camera stream representing parts of the transmitting participant of the virtual meeting;

based on a characteristic of the transmitting participant in the obtained live camera stream, select a non-realtime 3D model from a plurality of non-realtime 3D models that are general models that can be used for several different transmitting participants;

identify at least one body feature in the non-realtime 3D model;

identify at least one body feature in the live camera stream; and

generate a combined 3D representation of the transmitting participant based on both the non-realtime 3D model and the live camera stream, wherein generating the combined 3D representation further comprises, when the live camera stream is temporarily unavailable, combining the non-realtime 3D model with the most recently received live camera stream, and generating an appearance of the transmitting participant representing unavailability.

11 . The representation provider according to claim 10 , further comprising instructions that, when executed by the processor, cause the representation provider to:

generate the non-realtime 3D model based on camera data.

12 . The representation provider according to claim 10 , wherein generating the combined 3D representation comprises, for body features forming part of both the non-realtime 3D model and the live camera stream, assigning higher weights to data included in the live camera stream than data included in the non-realtime 3D model.

13 . The representation provider according to claim 10 , further comprising instructions that, when executed by the processor, cause the representation provider to:

transmit the combined 3D representation to a user device of the transmitting participant.

14 . The representation provider according to claim 13 , further comprising instructions that, when executed by the processor, cause the representation provider to:

receive a performance indication from the user device of the transmitting participant.

15 . A computer program product for providing a 3D, three-dimensional, representation of a transmitting participant in a virtual meeting, the computer program product comprising a non-transitory computer readable medium storing instructions which, when executed by a processor of a representation provider causes the representation provider to:

obtain a live camera stream representing parts of the transmitting participant of the virtual meeting;

based on a characteristic of the transmitting participant in the obtained live camera stream, select a non-realtime 3D model from a plurality of non-realtime 3D models that are general models that can be used for several different transmitting participants;

identify at least one body feature in the non-realtime 3D model;

identify at least one body feature in the live camera stream; and

generate a combined 3D representation of the transmitting participant based on both the non-realtime 3D model and the live camera stream, wherein generating the combined 3D representation further comprises, when the live camera stream is temporarily unavailable, combining the non-realtime 3D model with the most recently received live camera stream, and generating an appearance of the transmitting participant representing unavailability.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 26, 2023
From: EL ESSAILI, ALI; TYUDINA, NATALYA; EWERT, JOERG CHRISTIAN; MELANDER, OLA
To: TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)
Reel/Frame 063444/0880 →
Continuity (2)
Provisional Application 63116424 · Nov 20, 2020
Related Publication 20230396735A1 · Dec 7, 2023
References Cited (24)
US 12183035B1 · Chen · 2024 [cited by examiner]
US 20040133641A1 · McKinnon · 2004 [cited by examiner]
US 20060239186A1 · Wu · 2006 [cited by examiner]
US 20080024390A1 · Baker · 2008 [cited by examiner]
US 20080129844A1 · Cusack · 2008 [cited by examiner]
US 20140072270A1 · Goldberg · 2014 [cited by examiner]
US 20150042743A1 · Cullen · 2015 [cited by applicant]
US 20190297304A1 · Li · 2019 [cited by applicant]
US 20200204655A1 · Khalid · 2020 [cited by examiner]
US 20210125398A1 · Bradley · 2021 [cited by examiner]
US 20210209835A1 · Fonseka · 2021 [cited by examiner]
US 20230066958A1 · Condorovici · 2023 [cited by examiner]
CN 108513088A · 2018 [cited by applicant]
GB 2351425A · 2000 [cited by applicant]
JP 2000244886A · 2000 [cited by applicant]
JP 2016134730A · 2016 [cited by applicant]
JP 2020065229A · 2020 [cited by applicant]
Japanese Notice of Reasons for Rejection issued Aug. 20, 2024, for Japanese Patent Application No. 2023-530518, 4 pages (English translation). [cited by applicant]
International Search Report and Written Opinion of the International Searching Authority for PCT International Application No. PCT/EP2021/068756 dated Oct. 13, 2021. [cited by applicant]
Zollhöfer et al., “State of the Art on Monocular 3D Face Reconstruction, Tracking, and Applications,” Computer Graphics Forum : Journal of the European Association for Computer Graphics, vol. 37, No. 2, May 1, 2018 (May… [cited by applicant]
Lattas et al., “AvatarMe: Realistically Renderable 3D Facial Reconstruction “in-the-wild”,” arXiv:2003.13845v1 [cs.CV] Mar. 30, 2020, pp. 1-10. [cited by applicant]
Feng et al., “Joint 3D Face Reconstruction and Dense Alignment with Position Map Regression Network,” arXiv:1803.07835v1 [cs.CV] Mar. 21, 2018, pp. 1-18. [cited by applicant]
Taketomi et al., “Visual SLAM algorithms: a survey from 2010 to 2016,” IPSJ Transactions on Computer Vision and Applications (2017) 9:16, Review Paper, Open Access, pp. 1-11. [cited by applicant]
Richard et al., “Audio- and Gaze-driven Facial Animation of Codec Avatars,” arXiv:2008.05023v1 [cs.CV] Aug. 11, 2020, pp. 1-10. [cited by applicant]