IP Library › Granted Patent US 12,524,199
Granted Patent B2
US 12,524,199 · App. 18/291,183 · Granted Jan 13, 2026

Display device

Inventor: Woojin Choi (Seoul, KR)
Assignee: LG ELECTRONICS INC.
G06F3/167G06V10/56G06V10/82H04N21/42203H04N21/4312H04N21/472H04N21/482
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,524,199
App. No.
18/291,183
Granted
Jan 13, 2026
Kind
B2
Abstract

The present disclosure discloses a display device comprising: a communication unit; a display unit; a microphone; and a control unit which acquires screen image data by capturing a content provider screen including at least one piece of content outputted through the display unit upon receiving a wake-up word, inputs the screen image data to a screen recognition model corresponding to the content provider to detect at least one content region, outputs a labeling icon corresponding to each content region by using coordinate information of the content region corresponding to each of the at least one content regions detected by the screen recognition model, and when a user utterance corresponding to any one of the labeling icons is received, executes the content corresponding to the labeling icon.

Claims (36)

1 . A display device comprising:

a display unit configured to output a screen provided by a content provider;

a microphone; and

a controller configured to obtain screen image data by capturing a content provider screen including at least one content output through the display unit upon receiving a wake-up word, detect at least one content area by inputting the screen image data to a screen recognition model corresponding to the content provider, output a labelling icon corresponding to each of the at least one content area using coordinate information corresponding to each of the at least one content area detected by the screen recognition model, and execute content corresponding to the labelling icon upon receiving user utterance corresponding to any one of the labelling icon.

2 . The display device of claim 1 , wherein the coordinate information of the content area includes a content click coordinate and a content labelling coordinate.

3 . The display device of claim 2 , wherein the content click coordinate is a center coordinate, and the content labelling coordinate is an upper left coordinate of the content area.

4 . The display device of claim 1 , wherein, based on the screen image data being input, the screen recognition model lists up a contour candidate of the screen image data, aggregates the candidates to detect a content area, and detects a content area based on the detected region.

5 . The display device of claim 4 , wherein the controller separates the screen image data into R, G, and B channels and inputs each of separated channel images to the screen recognition model, and

the screen recognition model detects the content area of each of the R, G, and B channels, and aggregates content areas detected for the respective channels to detect the content area.

6 . The display device of claim 1 , wherein the screen recognition model includes an artificial neural network, and based on the screen image data being input, the screen recognition model is learned to output the coordinate information of the content area located in the screen.

7 . The display device of claim 1 , wherein the screen recognition model includes a first screen recognition model and a second screen recognition model, and

based on the screen image data being input, the first screen recognition model lists up a contour candidate of the screen image data and aggregates the candidates to detect a first content area,

the second screen recognition model includes an artificial neural network, and based on the screen image data being input, the screen recognition model detects a second content area located in the screen, and

the controller aggregates the first content area and the second content area to detect a content area.

8 . The display device of claim 1 , wherein the labelling icon includes at least one of object information displayed in the content area, text information displayed around the content area, and placement information.

9 . A method of operating a display device, the method comprising:

obtaining screen image data by capturing a content provider screen including at least one content output through a display unit upon receiving a wake-up word;

inputting the screen image data to a screen recognition model corresponding to the content provider;

detecting at least one content area using the screen recognition model;

obtaining coordinate information of a content area of the detected content area;

outputting a labelling icon corresponding to each of the at least one content area using the coordinate information of the content area corresponding to each of the at least one content area detected by the screen recognition model; and

upon receiving user utterance corresponding to any one of the labelling icon, executing content corresponding to the labelling icon.

10 . The method of claim 9 , wherein the coordinate information of the content area includes a content click coordinate and a content labelling coordinate.

11 . The method of claim 10 , wherein the content click coordinate is a center coordinate, and the content labelling coordinate is an upper left coordinate of the content area.

12 . The method of claim 9 , wherein the detecting of the at least one content area using the screen recognition model includes:

based on the screen image data being input, listing up a contour candidate of the screen image data, aggregating the candidates to detect a content area, and detecting a content area based on the detected region, by the screen recognition model.

13 . The method of claim 12 , wherein the detecting of the at least one content area using the screen recognition model includes:

separating the screen image data into R, G, and B channels and inputting each of separated channel images to the screen recognition model; and

detecting the content area of each of the R, G, and B channels, and aggregating content areas detected for the respective channels to detect the content area.

14 . The method of claim 9 , wherein the screen recognition model includes an artificial neural network, and based on the screen image data being input, the screen recognition model is learned to output the coordinate information of the content area located in the screen.

15 . The method of claim 9 , wherein the screen recognition model includes a first screen recognition model and a second screen recognition model,

the detecting of the at least one content area using the screen recognition model includes:

based on the screen image data being input, listing up a contour candidate of the screen image data and aggregating the candidates to detect a first content area, by the first screen recognition model;

based on the screen image data being input, the screen recognition model detects a second content area located in the screen, the second screen recognition model including an artificial neural network, and

aggregating the content area and the content area to detect a content area.

16 . The method of claim 9 , wherein the labelling icon includes at least one of object information displayed in the content area, text information displayed around the content area, and placement information.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 23, 2024
From: CHOI, WOOJIN
To: LG ELECTRONICS INC.
Reel/Frame 066215/0606 →
Continuity (1)
Related Publication 20240373080A1 · Nov 7, 2024
References Cited (22)
US 9277267B2 · Matsuda · 2016 [cited by examiner]
US 10353564B2 · Voutta · 2019 [cited by examiner]
US 12380340B1 · Mamut · 2025 [cited by examiner]
US 12400750B1 · Jin · 2025 [cited by examiner]
US 20130176244A1 · Yamamoto · 2013 [cited by examiner]
US 20160182577A1 · Lipman · 2016 [cited by examiner]
US 20170324794A1 · Jeong et al. · 2017 [cited by applicant]
US 20180114326A1 · Roblek et al. · 2018 [cited by applicant]
US 20180335908A1 · Kim · 2018 [cited by examiner]
US 20190050666A1 · Kim et al. · 2019 [cited by applicant]
US 20190371320A1 · Netzer · 2019 [cited by applicant]
US 20210117152A1 · Deisher · 2021 [cited by examiner]
US 20250265523A1 · Watanabe · 2025 [cited by examiner]
CN 109218526B · 2020 [cited by examiner]
CN 120144797A · 2025 [cited by examiner]
CN 120375383A · 2025 [cited by examiner]
KR 1020160091628 · 2016 [cited by applicant]
KR 1020170101076 · 2017 [cited by applicant]
KR 20180135074A · 2018 [cited by examiner]
KR 1020180135074 · 2018 [cited by applicant]
KR 102005034 · 2019 [cited by applicant]
PCT International Application No. PCT/KR2021/009581, International Search Report dated Apr. 20, 2022, 2 pages. [cited by applicant]