IP Library Granted Patent US 12,537,908
Granted Patent B2
US 12,537,908 · App. 18/735,626 · Granted Jan 27, 2026

Systems and methods for presence-aware repositioning and reframing in video conferencing

Inventor: Tao Chen (Palo Alto, CA)
Assignee: Adeia Guides Inc.
H04N5/2628G06T3/40G06T7/70G06V20/41G06T2207/10016G06T2207/20132G06T2207/30196
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,537,908
App. No.
18/735,626
Granted
Jan 27, 2026
Kind
B2
Abstract

Systems and methods are described herein for automatically reframing a video for a conference participant. Video of the participant is captured, and a first position of the participant is detected. An offset for the first position of the participant is then calculated to determine a relative distance from the center of the video frame. The captured video is modified based on the offset and is then presented in the video conference.

Claims (68)

1 . A method for automatically reframing video conference participants in a video stream of a video conference, the method comprising:

capturing video;

generating a pixel coordinate array for each respective frame of the video;

detecting within a first frame of the video, a plurality of participants, wherein the plurality of participants comprises a left-most participant and a right-most participant;

determining a left-most pixel area corresponding to the left-most participant and a right-most pixel area corresponding to the right-most participant, wherein the left-most pixel area and the right-most pixel area are both based on the pixel coordinate array;

calculating a group center based on the left-most pixel area and the right-most pixel area, wherein the group center is quantified using pixels from the pixel coordinate array;

modifying subsequent frames of the video based on the group center; and

presenting the modified subsequent frames of the video.

2 . The method of claim 1 , wherein modifying the subsequent frames of the video based on the group center further comprises:

translating a position of each participant of the plurality of participants to a new position corresponding to the group center of the first frame.

3 . The method of claim 2 , wherein translating the position of each participant of the plurality of participants to the new position corresponding to the group center of the first frame comprises:

calculating a translation vector; and

for each frame of the video, applying the translation vector to each pixel of a plurality of pixels of the respective frame that form an image of each participant.

4 . The method of claim 1 , wherein modifying the subsequent frames of the video based on the group center further comprises cropping the video so that the video is centered at the group center.

5 . The method of claim 1 , further comprising:

encoding a media stream including the video with the modified subsequent frames; and

transmitting, to a video conference server, the media stream.

6 . The method of claim 1 , further comprising:

encoding a media stream including the video and the group center;

modifying the video to incorporate the modified subsequent frames; and

transmitting, to a video conference server, the media stream.

7 . The method of claim 6 , wherein presenting the modified subsequent frames of the video further comprises:

retrieving, at the video conference server, from the media stream, the group center;

cropping, at the video conference server, the video based on the group center;

reencoding, at the video conference server, the cropped video in a second media stream; and

transmitting, from the video conference server, the second media stream to client devices associated with each participant in the video conference.

8 . The method of claim 1 , further comprising:

determining a first resolution of the video;

determining a second resolution to which to scale the video to fit in a video conference layout; and

scaling the video to the second resolution.

9 . The method of claim 1 , further comprising:

based on determining that a change in one or more of the left-most pixel area and the right-most pixel area is greater than a threshold amount, altering the group center.

10 . The method of claim 1 , wherein the group center is a midpoint between the positions of each detected participant.

11 . A system for automatically reframing video conference participants in a video stream of a video conference, the system comprising:

video capture circuitry configured to capture video;

input/output circuitry; and

control circuitry configured to:

generate a pixel coordinate array for each respective frame of the video;

detect within a first frame of the video, a plurality of participants, wherein the plurality of participants comprises a left-most participant and a right-most participant;

determine a left-most pixel area corresponding to the left-most participant and a right-most pixel area corresponding to the right-most participant, wherein the left-most pixel area and the right-most pixel area are both based on the pixel coordinate array;

calculate a group center based on the left-most pixel area and the right-most pixel area, wherein the group center is quantified using pixels from the pixel coordinate array;

modify subsequent frames of the video based on the group center; and

present the modified subsequent frames of the video.

12 . The system of claim 11 , wherein the control circuitry configured to modify the subsequent frames of the video based on the group center is further configured to:

translate a position of each participant of the plurality of participants to a new position corresponding to the group center of the first frame.

13 . The system of claim 12 , wherein the control circuitry configured to translate the position of each participant of the plurality of participants to the new position corresponding to the group center of the first frame is further configured to:

calculate a translation vector; and

for each frame of the video, apply the translation vector to each pixel of a plurality of pixels of the respective frame that form an image of each participant.

14 . The system of claim 11 , wherein the control circuitry configured to modify the subsequent frames of the video based on the group center is further configured to crop the video so that the video is centered at the group center.

15 . The system of claim 11 , wherein the control circuitry is further configured to:

encode a media stream including the video with the modified subsequent frames; and

transmit, to a video conference server, the media stream.

16 . The system of claim 11 , wherein the control circuitry is further configured to:

encode a media stream including the video and the group center;

modify the video to incorporate the modified subsequent frames; and

transmit, to a video conference server, the media stream.

17 . The system of claim 16 , wherein the control circuitry configured to present the modified subsequent frames of the video is further configured to:

retrieve, at the video conference server, from the media stream, the group center;

crop, at the video conference server, the video based on the group center;

reencode, at the video conference server, the cropped video in a second media stream; and

transmit, from the video conference server, the second media stream to client devices associated with each participant in the video conference.

18 . The system of claim 11 , wherein the control circuitry is further configured to:

determine a first resolution of the video;

determine a second resolution to which to scale the video to fit in a video conference layout; and

scale the video to the second resolution.

19 . The system of claim 11 , wherein the control circuitry is further configured to:

based on determining that a change in one or more of the left-most pixel area and the right-most pixel area is greater than a threshold amount, alter the group center.

20 . The system of claim 11 , wherein the group center is a midpoint between the positions of each detected participant.

Assignments (3)
SECURITY INTEREST Recorded May 28, 2025
From: ADEIA INC. (F/K/A XPERI HOLDING CORPORATION); ADEIA HOLDINGS INC.; ADEIA MEDIA HOLDINGS INC.; ADEIA IMAGING LLC; ADEIA MEDIA LLC; ADEIA MEDIA SOLUTIONS INC.; ADEIA SEMICONDUCTOR BONDING TECHNOLOGIES INC.; ADEIA TECHNOLOGIES INC.; ADEIA GUIDES INC.; ADEIA SOLUTIONS LLC; ADEIA SEMICONDUCTOR ADVANCED TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR SOLUTIONS LLC; ADEIA SEMICONDUCTOR INTELLECTUAL PROPERTY LLC; ADEIA SEMICONDUCTOR TECHNOLOGIES LLC; ADEIA PUBLISHING INC.
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 071454/0343 →
CHANGE OF NAME Recorded Oct 4, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069113/0413 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 6, 2024
From: CHEN, TAO
To: ROVI GUIDES, INC.
Reel/Frame 067646/0547 →
Continuity (2)
Continuation 17828223 · May 31, 2022
Related Publication 20240323312A1 · Sep 26, 2024
References Cited (14)
US 5986703A · O'Mahony · 1999 [cited by examiner]
US 9253442B1 · Pauli · 2016 [cited by applicant]
US 20060215765A1 · Hwang et al. · 2006 [cited by applicant]
US 20110115876A1 · Khot · 2011 [cited by examiner]
US 20110310214A1 · Saleh · 2011 [cited by examiner]
US 20120027085A1 · Amon et al. · 2012 [cited by applicant]
US 20150288926A1 · Glass · 2015 [cited by examiner]
US 20160267631A1 · Shen · 2016 [cited by examiner]
US 20190215464A1 · Kumar · 2019 [cited by examiner]
US 20200342652A1 · Rowell · 2020 [cited by examiner]
US 20220072433A1 · Wang · 2022 [cited by examiner]
US 20230230210A1 · Powell · 2023 [cited by examiner]
US 20230247069A1 · Khire · 2023 [cited by examiner]
US 20230388444A1 · Chen · 2023 [cited by applicant]