IP Library › Granted Patent US 12,657,670
Granted Patent B2
US 12,657,670 · App. 18/395,834 · Granted Jun 16, 2026

Key point smoothing method based on frame up-sampling

Inventors: Young Han Lee (Seongnam-si, KR); Choong Sang Cho (Seongnam-si, KR); Tae Woo Kim (Seongnam-si, KR)
Assignee: Korea Electronics Technology Institute
G06T5/70G06T3/40G06V40/168G10L25/57G06T2207/10016G06T2207/20032G06T2207/20081G06T2207/30201G06V20/46
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,657,670
App. No.
18/395,834
Granted
Jun 16, 2026
Kind
B2
Abstract

There is provided a frame up-sampling-based key point smoothing method. The key point smoothing method according to an embodiment up-samples frames based on which key points are extracted, and smooths the up-sampled frames. Accordingly, key points are generated by smoothing key point extraction frames after up-sampling, so that a time-series error occurring when key points are extracted can be reduced and time-series stability can be enhanced, and quality of an application service provided subsequently can be improved.

Claims (25)

1 . A key point generation method comprising:

extracting facial key points on a frame basis;

temporally up-sampling key point extraction frames based on which the facial key points are extracted on a time axis; and

smoothing the temporally up-sampled key point extraction frames on the time axis,

wherein the smoothing comprises smoothing a predetermined number of frames at predetermined intervals, wherein the predetermined interval is determined based on a frame rate before up-sampling, wherein the predetermined number is determined based on an increase rate of a frame rate caused by up-sampling, and wherein the method includes reducing a time-series error of the facial key points based on a result of the smoothing.

2 . The key point generation method of claim 1 , further comprising acquiring a speech signal,

wherein the extracting comprises extracting facial key points from the acquired speech signal on a frame basis.

3 . The key point generation method of claim 2 , wherein the extracting comprises inputting a speech signal to a machine learning model that is trained to extract facial key points from a speech signal, and extracting the facial key points.

4 . The key point generation method of claim 1 , further comprising providing an application service by using the smoothed key point extraction frames.

5 . A key point generation system comprising:

one or more processors configured to:

extract facial key points on a frame basis;

temporally up-sample key point extraction frames based on which the facial key points are extracted on a time axis; and

smooth the temporally up-sampled key point extraction frames on the time axis,

wherein the smoothing comprises smoothing a predetermined number of frames at predetermined intervals, wherein the predetermined interval is determined based on a frame rate before up-sampling, wherein the predetermined number is determined based on an increase rate of a frame rate caused by up-sampling, and wherein the one or more processors are configured to reduce a time-series error of the facial key points based on a result of the smoothing.

6 . The system of claim 5 , wherein the one or more processors are further configured to acquire a speech signal and extract facial key points from the acquired speech signal on a frame basis.

7 . The system of claim 6 , wherein, for the extracting, the one or more processors are configured to input a speech signal to a machine learning model that is trained to extract facial key points from a speech signal, and extract the facial key points.

8 . The system of claim 5 , wherein the one or more processors are configured to provide an application service by using the smoothed key point extraction frames.

9 . A key point smoothing method comprising:

temporally up-sampling key point extraction frames based on which facial key points are extracted on a frame basis on a time axis; and

smoothing the temporally up-sampled key point extraction frames on the time axis,

wherein the smoothing comprises smoothing a predetermined number of frames at predetermined intervals,

wherein the predetermined interval is determined based on a frame rate before up-sampling,

wherein the predetermined number is determined based on an increase rate of a frame rate caused by up-sampling, and

wherein the method includes reducing a time-series error of the facial key points based on a result of the smoothing.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 26, 2023
From: LEE, YOUNG HAN; CHO, CHOONG SANG; KIM, TAE WOO
To: KOREA ELECTRONICS TECHNOLOGY INSTITUTE
Reel/Frame 065951/0291 →
Priority Claims (1)
KR 10-2023-0178605 · Dec 11, 2023 · national
Continuity (1)
Related Publication 20250191140A1 · Jun 12, 2025
References Cited (13)
US 6330023B1 · Chen · 2001 [cited by examiner]
US 9794092B1 · Isautier · 2017 [cited by examiner]
US 12184923B2 · Zhang · 2024 [cited by examiner]
US 20090022226A1 · Bang · 2009 [cited by examiner]
US 20190178814A1 · Nakano · 2019 [cited by examiner]
US 20200302969A1 · Boyd · 2020 [cited by examiner]
US 20210295016A1 · Yan · 2021 [cited by examiner]
US 20220222968A1 · Cower · 2022 [cited by examiner]
US 20230147584A1 · Asgekar · 2023 [cited by examiner]
US 20250131599A1 · Chen · 2025 [cited by examiner]
KR 1020210140762A · 2021 [cited by applicant]
Korean Office Action issued on Apr. 1, 2025, in corresponding Korean Patent Application No. 10-2023-0178605. (3pages in English, 5pages in Korean). [cited by applicant]
Szeto, Ryan, et al. “HyperCon: Image-to-video model transfer for video-to-video translation tasks.” Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision, 2021. (pp. 3079-3088). [cited by applicant]