IP Library › Granted Patent US 12,288,404
Granted Patent B2
US 12,288,404 · App. 17/476,251 · Granted Apr 29, 2025

Resolution upscaling for event detection

Inventors: Nathan Otterness (Mebane, NC); Jonathan White (Fort Collins, CO); Dave Clark (Cary, NC); Jim van Welzen (Sandy, UT)
Assignee: Nvidia Corporation
G06V20/635A63F13/53G06T3/4007G06V20/41G06V20/46G06V30/153G06V20/44
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,288,404
App. No.
17/476,251
Granted
Apr 29, 2025
Kind
B2
Abstract

A game-agnostic event detector can be used to automatically identify game events. Game-specific configuration data can be used to specify types of pre-processing to be performed on media for a game session, as well as types of detectors to be used to detect events for the game. Event data for detected events can be written to an event log in a form that is both human- and process-readable. The event data can be used for various purposes, such as to generate highlight videos or provide player performance feedback. The event data may be determined based upon output from detectors such as optical character recognition (OCR) engines, and the regions may be upscaled and binarized before OCR processing.

Claims (56)

1. A computer-implemented method, comprising:

receiving video data corresponding to a session;

determining an event region in a frame of the video data, the event region including one or more objects associated with a type of event;

upscaling image data for the event region to obtain a higher graphical resolution for the event region;

processing the upscaled image data to recognize the one or more objects present in the event region;

determining, using a neural network, an inference of an occurrence of an event, the neural network determining the inference based on two or more correlated event cues and the one or more objects recognized from the event region in the higher graphical resolution; and

providing event data corresponding to the determined event.

2. The computer-implemented method of claim 1 , further comprising:

performing pre-processing of the image data before processing using optical character recognition (OCR), the pre-processing including binarization of the upscaled image data to remove background values from pixels of the event region.

3. The computer-implemented method of claim 1 , further comprising:

determining a size of the event region, the size including at least a minimum amount of pixel padding around the one or more objects.

4. The computer-implemented method of claim 1 , further comprising:

determining an upscaling algorithm to use for upscaling the image data, the upscaling algorithm including at least one of a bi-cubic, bilinear, edge-based, or fractal upscaling algorithm.

5. The computer-implemented method of claim 1 , further comprising:

determining the event based at least in part upon a current state, or change in state, of the one or more objects in the region.

6. The computer-implemented method of claim 1 , further comprising:

detecting the region using a pattern recognition algorithm.

7. The computer-implemented method of claim 1 , wherein the event region corresponds to a heads-up display (HUD) presented over rendered content.

8. The computer-implemented method of claim 1 , further comprising:

analyzing a periodic subset of frames of a video sequence to determine a plurality of events in the session.

9. The computer-implemented method of claim 1 , further comprising:

performing pre-processing of the image data, the pre-processing including at least one of filtering, stretching, warping, perspective correction, noise removal, color space transform, color isolation, or value thresholding.

10. The computer-implemented method of claim 1 , further comprising:

receiving, from an OCR engine, at least one of the two or more correlated event cues; and

utilizing at least one cue-to-event translation algorithm to determine the event data based, at least in part, upon the two or more correlated event cues.

11. The computer-implemented method of claim 1 , further comprising:

receiving the video data in the form of a file or stream during, or after, the session.

12. A computer-implemented method, comprising:

determining, in a frame of video data, an object region including one or more objects;

upscaling image data for the object region to obtain a higher graphical resolution for the object region than for the frame of video data;

performing pre-processing of the upscaled image data, the pre-processing including binarization of the upscaled image data to remove background values from pixels of the object region;

processing the upscaled and pre-processed image data to recognize the one or more objects present in the object region; and

providing content relating to the recognized one or more objects based at least in part on a neural network determination of an event, the neural network to compute an inference of an occurrence of the event, the neural network determining the inference based on two or more correlated event cues and the one or more objects recognized from the event region in the higher graphical resolution.

13. The computer-implemented method of claim 12 , further comprising:

determining a size of the object region, the size including at least a minimum amount of padding around the one or more objects.

14. The computer-implemented method of claim 12 , further comprising:

determining an upscaling algorithm to use for upscaling the image data, the upscaling algorithm including at least one of a bi-cubic, bilinear, edge-based, or fractal upscaling algorithm.

15. The computer-implemented method of claim 12 , further comprising:

determining the content to be provided based at least in part upon a current state, or change in state, of the one or more objects in the object region.

16. The computer-implemented method of claim 12 , further comprising:

detecting the object region using a pattern recognition algorithm.

17. A system comprising:

one or more processors; and

memory including instructions that, when executed by the one or more processors, cause the system to:

receive video data corresponding to a session for a user;

determine an event region in a frame of the video data, the event region including one or more objects associated with a type of event;

upscale image data for the event region to obtain a higher graphical resolution for the event region;

process the upscaled image data to recognize the one or more objects present in the event region;

determine, using a neural network, an inference of an occurrence of an event the neural network determining the inference based on two or more correlated event cues and the one or more objects recognized from the event region in the higher graphical resolution; and

provide event data corresponding to the determined event.

18. The system of claim 17 , wherein the instructions when executed further cause the system to:

perform pre-processing of the image data before processing using optical character recognition (OCR), the pre-processing including binarization of the upscaled image data to remove background values from pixels of the event region.

19. The system of claim 17 , wherein the instructions when executed further cause the system to:

determine a size of the event region, the size including at least a minimum amount of padding around the one or more objects.

20. The system of claim 17 , wherein the instructions when executed further cause the system to:

determine an upscaling algorithm to use for upscaling the image data, the upscaling algorithm including at least one of a bi-cubic, bilinear, edge-based, or fractal upscaling algorithm.

Continuity (2)
Continuation 16747143 · Jan 20, 2020
Related Publication 20220005156A1 · Jan 6, 2022
References Cited (49)
US 8944234B1 · Csulits · 2015 [cited by applicant]
US 9007189B1 · Curtis et al. · 2015 [cited by applicant]
US 10737183B1 · Sharma · 2020 [cited by applicant]
US 11170471B2 · Otterness · 2021 [cited by examiner]
US 20090088233A1 · O'Rourke · 2009 [cited by examiner]
US 20140212853A1 · Divakaran · 2014 [cited by applicant]
US 20150222239A1 · Zhang · 2015 [cited by applicant]
US 20170065888A1 · Cheng · 2017 [cited by examiner]
US 20170113143A1 · Marr · 2017 [cited by applicant]
US 20170128843A1 · Versaci · 2017 [cited by applicant]
US 20170136367A1 · Watari · 2017 [cited by examiner]
US 20170140570A1 · Leibel · 2017 [cited by applicant]
US 20170228600A1 · Syed · 2017 [cited by examiner]
US 20170282077A1 · De La Cruz · 2017 [cited by applicant]
US 20170294081A1 · Washington · 2017 [cited by applicant]
US 20180338120A1 · Lemberger et al. · 2018 [cited by applicant]
US 20190028306A1 · Namise · 2019 [cited by applicant]
US 20190099668A1 · Aliakseyeu · 2019 [cited by applicant]
US 20190111347A1 · Rimon · 2019 [cited by applicant]
US 20190351335A1 · Yong et al. · 2019 [cited by applicant]
US 20200030697A1 · Mueller · 2020 [cited by applicant]
US 20200090371A1 · Hu · 2020 [cited by applicant]
US 20200126183A1 · Zhu · 2020 [cited by examiner]
US 20200175303A1 · Bhat et al. · 2020 [cited by applicant]
US 20200211601A1 · Shanmuga Vadivel · 2020 [cited by examiner]
US 20200346121A1 · Beaumont · 2020 [cited by applicant]
CN 108769821A · 2018 [cited by applicant]
CN 109076678A · 2018 [cited by applicant]
CN 110443238A · 2019 [cited by applicant]
JP 2002157374A · 2002 [cited by applicant]
Non-Final Office Action issued in U.S. Appl. No. 16/745,881 dated Jun. 1, 2021. [cited by applicant]
International Search Report and Written Opinion issued in PCT Application No. PCT/US2021 /013632 dated Mar. 30, 2021. [cited by applicant]
Notice of Allowance issued in U.S. Appl. No. 16/669,939, dated Jul. 8, 2021. [cited by applicant]
Non-Final Office Action issued in U.S. Appl. No. 16/747,143, dated Mar. 11, 2021. [cited by applicant]
Notice of Allowance issued in U.S. Appl. No. 16/747,143, dated Jul. 21, 2021. [cited by applicant]
International Search Report and Written Opinion issued in PCT Application No. PCT/US2021/013973 dated Apr. 15, 2021. [cited by applicant]
Yun-Gyung Cheong et al: “Automatically Generating Surrmary Visualizations from Game Logs”, Proceedings of the Fourth Artificial Intelligence and Interactive Digital Entertainment Conference, Oct. 24, 2008 (Oct. 24, 2008… [cited by applicant]
Shih Huang-Chia: “A Survey of Content-Aware Video Analysis for Sports”, IEEE Transactions on Circuits and Systems for Video Technology, Institute of Electrical and Electronics Engineers, US, vol. 28, No. 5, May 31, 2018… [cited by applicant]
Guo Jinlin et al: “Localization and Recognition of the Scoreboard in Sports Video Based on SIFT Point Matching”, Jan. 5, 2011 (Jan. 5, 2011), ICIAP: International Conference on Image Analysis and Processing, 17th Intern… [cited by applicant]
Final Office Action issued in U.S. Appl. No. 16/745,881, dated Dec. 6, 2021. [cited by applicant]
Non-Final Office Action issued in U.S. Appl. No. 16/745,881, dated May 27, 2022. [cited by applicant]
Final Office Action issued in U.S. Appl. No. 16/745,881, dated Oct. 13, 2022. [cited by applicant]
Notice of Allowance issued in U.S. Appl. No. 16/745,881, dated Feb. 23, 2023. [cited by applicant]
Non-Final Office Action issued in U.S. Appl. No. 16/669,939, dated Nov. 30, 2020. [cited by applicant]
Non-Final Office Action issued in U.S. Appl. No. 17/495,643, dated Sep. 29, 2022. [cited by applicant]
Corrected Notice of Allowance Action issued in U.S. Appl. No. 16/747,143, dated Jul. 30, 2021. [cited by applicant]
Final Office Action issued in U.S. Appl. No. 17/495,643, dated Jun. 1, 202. [cited by applicant]
Notice of Allowance issued in U.S. Appl. No. 17/495,643, dated Jul. 19, 2023. [cited by applicant]
“Theory and Application of Intelligent Vehicles”, Beijing Institute of Technology Press, ISBN 978-7-5682-5966-8, Jul. 2018, 3 pages. [cited by applicant]