IP Library Granted Patent US 12,277,799
Granted Patent B2
US 12,277,799 · App. 17/804,100 · Granted Apr 15, 2025

Video image composition method and electronic device

Inventors: Yen-Chou Chen (Taipei, TW); Chui-Pang Chiu (New Taipei, TW); Che-Chia Ho (New Taipei, TW)
Assignee: AmTRAN Technology Co., Ltd.
G06V40/166G06V40/172G10L15/25
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,277,799
App. No.
17/804,100
Granted
Apr 15, 2025
Kind
B2
Abstract

The present disclosure provides a video image composition method including the following steps. A priority level list is obtained, and the priority level list includes multiple priority levels of multiple person identities. Multiple video streams are received. Multiple identity labels corresponding to human face frame images from the video streams are determined. The multiple display levels of the human face frame images are determined according to the identity labels and priority level list. A part of the human face frame images being in speaking status are detected. At least one of the part of the human face frame images being in speaking status is constituted as a main display area of a video image, according to the display levels.

Claims (39)

1. A video image composition method, comprising:

receiving a priority level list, wherein the priority level list comprises a plurality of priority levels of a plurality of person identities;

receiving a plurality of video streams;

identifying a plurality of identity labels corresponding to a plurality of human face frame images in the video streams;

determining a plurality of display levels corresponding to the human face frame images, according to the identity labels and the priority level list;

detecting a part of the human face frame images being in speaking status;

constituting at least one of the part of the human face frame images being in speaking status as a main display area of a video image, according to the display levels; and

in a moderator mode, determining a first human face frame image from the human face frame images being speaking, wherein the first human face frame image is corresponding to a first identity label comprising a highest display priority order of the display levels.

2. The video image composition method of claim 1 , further comprising:

in the moderator mode, configuring the first human face frame image to the main display area of the video image.

3. The video image composition method of claim 1 , further comprising:

determining a second human face frame image corresponding to a second identity label according to an indication from a person corresponding to the first identity label, and wherein in response to the indication, decomposing the main display area into a first split area and a second split area and starting a question-and-answer mode.

4. The video image composition method of claim 3 , further comprising:

in the question-and-answer mode, configuring the first human face frame image to the first split area, and configuring the second human face frame image to the second split area.

5. The video image composition method of claim 4 , further comprising:

in response to end of the question-and-answer mode, switching the question-and-answer mode back to the moderator mode, to combine the first split area and the second split area into the main display area.

6. The video image composition method of claim 3 , wherein the indication from the person corresponding to the first identity label is a sound source signal received by a sound station.

7. The video image composition method of claim 6 , wherein the sound source signal comprises the second identity label and a keyword of the question-and-answer mode.

8. An electronic device, comprising:

a memory device; and

a processing circuit, configured to:

receiving a priority level list, wherein the priority level list comprises a plurality of priority levels of a plurality of person identities;

receiving a plurality of video streams;

identifying a plurality of identity labels corresponding to a plurality of human face frame images in the video streams;

searching the priority level list for the priority levels according to the identity labels to determine a plurality of display levels corresponding to the human face frame images;

detecting a part of the human face frame images being in speaking status; and

constituting at least one of the part of the human face frame images being in speaking status as a main display area of a video image, according to the display levels.

9. The electronic device of claim 8 , wherein the processing circuit is configured to:

in a moderator mode, determining a first human face frame image from the human face frame images being speaking, and wherein the first human face frame image is corresponding to a first identity label comprising a highest display level of the display levels.

10. The electronic device of claim 9 , wherein the processing circuit is configured to:

in the moderator mode, configuring the first human face frame image to the main display area of the video image.

11. The electronic device of claim 10 , wherein the processing circuit is configured to:

determining a second human face frame image corresponding to a second identity label according to an indication from a person corresponding to the first identity label, and wherein in response to the indication, decomposing the main display area into a first split area and a second split area and starting a question-and-answer mode.

12. The electronic device of claim 11 , wherein the processing circuit is configured to:

in the question-and-answer mode, configuring the first human face frame image to the first split area, and configuring the second human face frame image to the second split area.

13. The electronic device of claim 12 , wherein the processing circuit is configured to:

in response to end of the question-and-answer mode, switching the question-and-answer mode back to the moderator mode, to combine the first split area and the second split area into the main display area.

14. The electronic device of claim 11 , wherein the indication from the person corresponding to the first identity label is a sound source signal received by a sound station.

15. The electronic device of claim 14 , wherein the sound source signal comprises the second identity label and a keyword of the question-and-answer mode.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 26, 2022
From: CHEN, YEN-CHOU; CHIU, CHUI-PANG; HO, CHE-CHIA
To: AMTRAN TECHNOLOGY CO., LTD.
Reel/Frame 060033/0153 →
Priority Claims (1)
TW 111103019 · Jan 24, 2022 · national
Continuity (1)
Related Publication 20230237838A1 · Jul 27, 2023
References Cited (18)
US 6567775B1 · Maali · 2003 [cited by examiner]
US 10304458B1 · Woo · 2019 [cited by examiner]
US 20130216206A1 · Dubin · 2013 [cited by examiner]
US 20150189233A1 · Carpenter · 2015 [cited by examiner]
US 20160371534A1 · Koul · 2016 [cited by examiner]
US 20170244930A1 · Faulkner · 2017 [cited by examiner]
US 20180376108A1 · Bright-Thomas · 2018 [cited by examiner]
US 20190222892A1 · Faulkner · 2019 [cited by examiner]
US 20190273767A1 · Nelson · 2019 [cited by examiner]
US 20190379839A1 · Hellerud · 2019 [cited by examiner]
US 20200026729A1 · Narasimha · 2020 [cited by examiner]
US 20200186649A1 · Zheng · 2020 [cited by examiner]
US 20200344278A1 · Mackell · 2020 [cited by examiner]
US 20210185276A1 · Peters · 2021 [cited by examiner]
US 20210399911A1 · Jorasch · 2021 [cited by examiner]
CN 109413359A · 2019 [cited by applicant]
CN 109819195B · 2021 [cited by examiner]
TW 201901527A · 2019 [cited by applicant]