IP Library Granted Patent US 12,355,992
Granted Patent B2
US 12,355,992 · App. 17/895,561 · Granted Jul 8, 2025

Method and apparatus for artificial intelligence downscaling and upscaling during video conference

Inventors: Chaeeun Lee (Suwon-si, KR); Jaehwan Kim (Suwon-si, KR); Youngo Park (Suwon-si, KR); Jongseok Lee (Suwon-si, KR); Kwangpyo Choi (Suwon-si, KR)
Assignee: SAMSUNG ELECTRONICS CO., LTD.
H04N19/30H04N19/136H04N19/164H04N19/436
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,355,992
App. No.
17/895,561
Granted
Jul 8, 2025
Kind
B2
Abstract

Provided is an electronic device configured to participate in video conference by using artificial intelligence (AI), the electronic device including a display and a processor configured to execute one or more instructions. The processor configured to obtain, from a server, image data generated by first encoding first image related to another electronic device participating in video conference, and AI data related to AI downscaling from original image to first image, obtain second image corresponding to first image by performing first decoding on image data, determine whether to perform AI upscaling on second image, based on importance of the other electronic device, based on determining to perform AI upscaling, obtain third image by performing AI upscaling on second image through an upscaling deep neural network and provide third image to display, and based on determining not to perform AI upscaling, provide second image to display.

Claims (36)

1. An electronic device configured to be in a video conference by using artificial intelligence (AI), the electronic device comprising:

a display; and

a processor configured to execute one or more instructions stored in the electronic device,

wherein the processor is configured to execute the one or more instructions to:

obtain, from a server, first image data generated as a result of first encoding a first image related to a first electronic device in the video conference, and first AI data related to AI downscaling from a first original image to the first image;

obtain, from the server, second image data generated as a result of the first encoding a second image related to a second electronic device in the video conference, and second AI data related to AI downscaling from a second original image to the second image;

obtain a third image corresponding to the first image by performing first decoding on the first image data, and fourth image corresponding to the second image by performing the first decoding on the second image data;

based on an importance of the first electronic device indicating that a user of the first electronic device is a presenter, obtain a fifth image by performing AI upscaling on the third image through an upscaling deep neural network (DNN), and provide the fifth image to the display; and based on an importance of the second electronic device indicating that a user of the second electronic device is a listener, obtain a sixth image by performing AI downscaling on the fourth image through an downscaling deep neural network (DNN), and provide the sixth image to the display.

2. The electronic device of claim 1 , wherein the importance of the first electronic device is checked from the first AI data, and the importance of the second electronic device is checked from the second AI data.

3. The electronic device of claim 1 , wherein the first image is an image to which the server AI-downscales the first original image or an image to which the first electronic device AI-downscales the first original image.

4. The electronic device of claim 1 , wherein

based on the importance of the first electronic device being changed, during the video conference, to indicate that the user of the first electronic device is the listener, it is determined that the AI upscaling is not performed.

5. The electronic device of claim 1 , wherein importance of the electronic device that establishes the video conference is initially set as the presenter.

6. A video conference image processing method performed by an electronic device in a video conference by using artificial intelligence (AI), the video conference image processing method comprising:

obtaining, from a server, first image data generated as a result of first encoding a first image related to a first electronic device in the video conference, and first AI data related to AI downscaling from a first original image to the first image;

obtaining, from the server, second image data generated as a result of the first encoding a second image related to a second electronic device in the video conference, and second AI data related to AI downscaling from a second original image to the second image;

obtaining a third image corresponding to the first image by performing first decoding on the first image data, and fourth image corresponding to the second image by performing the first decoding on the second image data;

based on an importance of the first electronic device indicating that a user of the first electronic device is a presenter, obtaining a fifth image by performing AI upscaling on the third image through an upscaling deep neural network (DNN), and provide the fifth image to a display; and

based on an importance of the second electronic device indicating that a user of the second electronic device is a listener, obtaining a sixth image by performing AI downscaling on the fourth image through an downscaling deep neural network (DNN), and provide the sixth image to the display.

7. A server managing a video conference by using artificial intelligence (AI), the server comprising a processor configured to execute one or more instructions stored in the server,

wherein the processor is configured to execute the one or more instructions to:

obtain, from a first electronic device in the video conference, first image data generated as a result of first encoding a first image, and first AI data related to the AI downscaling from a first original image to the first image;

obtain, from a second electronic device in the video conference, second image data generated as a result of the first encoding a second image, and second AI data related to AI downscaling from a second original image to the second image

obtain a third image corresponding to the first image by performing first decoding on the first image data, and fourth image corresponding to the second image by performing the first decoding on the second image data;

based on an importance of the first electronic device indicating that a user of the first electronic device is a presenter, obtain a fifth image by performing AI upscaling on the third image through an upscaling deep neural network (DNN), and transmit the fifth image to a third electronic device; and

based on an importance of the second electronic device indicating that a user of the second electronic device is a listener, obtain a sixth image by performing AI downscaling on the fourth image through a downscaling deep neural network (DNN), and transmit the sixth image to the third electronic device.

8. The server of claim 7 , wherein the importance of the first electronic device is checked from the first AI data related to the AI downscaling.

9. The server of claim 7 , wherein the first electronic device is configured to support AI downscaling.

10. The server of claim 7 , wherein the second electronic device is configured not to support AI upscaling.

11. The server of claim 7 , wherein the processor is further configured to execute the one or more instructions to:

obtain third image data by performing first encoding on a third original image from a fourth electronic device;

obtain the third original image by performing first decoding on the third image data;

based on importance of the fourth electronic device indicating that a user of the fourth electronic device is the listener, obtain the first image by performing AI downscaling on the first original image by using the downscaling DNN, and transmit fourth image data obtained by performing first encoding on the first image to the third electronic device; and

based on the importance of the fourth electronic device indicating that the user of the first electronic device is the presenter, transmit the third image data to the third electronic device.

12. The server of claim 11 , wherein the fourth electronic device is configured not to support AI downscaling.

13. The server of claim 11 , wherein, based on the third electronic device being configured to support AI upscaling, the processor is further configured to obtain the first image by performing AI downscaling on the first original image by using the downscaling DNN, and transmit, to the third electronic device, the third image data obtained by performing first encoding on the first image and the first AI data related to the AI downscaling.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 25, 2022
From: LEE, CHAEEUN; KIM, JAEHWAN; PARK, YOUNGO; LEE, JONGSEOK; CHOI, KWANGPYO
To: SAMSUNG ELECTRONICS CO., LTD.
Reel/Frame 060902/0215 →
Priority Claims (1)
KR 10-2021-0112653 · Aug 25, 2021 · national
Continuity (2)
Continuation PCTKR2022012646 · Aug 24, 2022
Related Publication 20230085530A1 · Mar 16, 2023
References Cited (36)
US 10817990B1 · Yang · 2020 [cited by examiner]
US 10825139B2 · Kim et al. · 2020 [cited by applicant]
US 10825205B2 · Kim et al. · 2020 [cited by applicant]
US 11200702B2 · Dinh et al. · 2021 [cited by applicant]
US 11349893B1 · Viswanathan · 2022 [cited by examiner]
US 11893067B1 · Solh · 2024 [cited by examiner]
US 12069121B1 · Garcia · 2024 [cited by examiner]
US 20150009276A1 · Maxwell · 2015 [cited by applicant]
US 20160275354A1 · Andalo et al. · 2016 [cited by applicant]
US 20180070054A1 · Miyamoto · 2018 [cited by applicant]
US 20190373293A1 · Bortman et al. · 2019 [cited by applicant]
US 20200099889A1 · Sugihara · 2020 [cited by applicant]
US 20200267349A1 · Garrido et al. · 2020 [cited by applicant]
US 20210056666A1 · Park et al. · 2021 [cited by applicant]
US 20210166343A1 · Kim et al. · 2021 [cited by applicant]
US 20210174552A1 · Lee et al. · 2021 [cited by applicant]
US 20210211739A1 · Andreopoulos · 2021 [cited by examiner]
US 20210390662A1 · Kim et al. · 2021 [cited by applicant]
US 20220138904A1 · Kim et al. · 2022 [cited by applicant]
US 20220229626A1 · Lysdal · 2022 [cited by examiner]
US 20240089136A1 · Tandra · 2024 [cited by examiner]
US 20240098123A1 · Gan · 2024 [cited by examiner]
JP 2001313938A · 2001 [cited by applicant]
JP 202053741A · 2020 [cited by applicant]
KR 1020180082672A · 2018 [cited by applicant]
KR 1020200044662A · 2020 [cited by applicant]
KR 1020200044666A · 2020 [cited by applicant]
KR 1020200044667A · 2020 [cited by applicant]
KR 1020200140368A · 2020 [cited by applicant]
KR 1020210050186A · 2021 [cited by applicant]
KR 1020210154700A · 2021 [cited by applicant]
WO 2016151974A1 · 2016 [cited by applicant]
WO 2020080873A1 · 2020 [cited by applicant]
Jiang, Feng, et al. “An end-to-end compression framework based on convolutional neural networks.” (Year: 2017). [cited by examiner]
International Search Report (PCT/ISA/220 and PCT/ISA/210) and Written Opinion (PCT/ISA/237) dated Nov. 28, 2022 by the International Searching Authority in International Application No. PCT/KR2022/012646. [cited by applicant]
Extended European Search Report dated May 17, 2024, issued by the European Patent Office in European Application No. 22861712.2. [cited by applicant]