IP Library Granted Patent US 11,861,850
Granted Patent B2
US 11,861,850 · App. 18/171,066 · Granted Jan 2, 2024

System and method for player reidentification in broadcast video

Inventors: Long Sha (Chicago, IL); Sujoy Ganguly (Chicago, IL); Xinyu Wei (Melbourne, AU); Patrick Joseph Lucey (Chicago, IL); Aditya Cherukumudi (London, GB)
Assignee: STATS LLC
G06T7/20G06F18/214G06F18/2135G06F18/22G06F18/232G06F18/2413G06N3/08G06T7/70G06T7/73G06T7/80G06T7/97G06V10/454G06V10/764G06V10/82G06V20/42G06V20/46G06V20/48G06V20/49G06V40/20H04N21/44008G06T2207/10016G06T2207/20081G06T2207/20084G06T2207/30221G06T2207/30244G06V20/44
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,861,850
App. No.
18/171,066
Granted
Jan 2, 2024
Kind
B2
Abstract

A system and method of re-identifying players in a broadcast video feed are provided herein. A computing system retrieves a broadcast video feed for a sporting event. The broadcast video feed includes a plurality of video frames. The computing system generates a plurality of tracks based on the plurality of video frames. Each track includes a plurality of image patches associated with at least one player. Each image patch of the plurality of image patches is a subset of the corresponding frame of the plurality of video frames. For each track, the computing system generates a gallery of image patches. A jersey number of each player is visible in each image patch of the gallery. The computing system matches, via a convolutional autoencoder, tracks across galleries. The computing system measures, via a neural network, a similarity score for each matched track and associates two tracks based on the measured similarity.

Claims (58)

1. A method comprising:

identifying, by a computing system, broadcast video feed of a game, the broadcast video feed captured by a broadcast camera;

identifying, by the computing system from the broadcast video feed, a first track corresponding to a first player, the first track comprising trajectory information corresponding to the first player up to a first time, wherein after the first time, the first player is not a field of view of the broadcast camera;

identifying, by the computing system from the broadcast video feed, a second track corresponding to a second player, the second track comprising second trajectory information corresponding to the second player up to a second time, the second time after the first time;

learning, by the computing system, a first set of visual attributes corresponding to the first player in the first track;

learning, by the computing system, a second set of visual attributes corresponding to the second player in the second track;

determining, by the computing system, that the first player and the second player are the same player by measuring a similarity between the first set of visual attributes and the second set of visual attributes; and

inferring, by the computing system, movement of the first player between the first time and the second time, wherein the first player is not in the field of view of the broadcast camera between the first time and the second time.

2. The method of claim 1 , wherein learning, by the computing system, the first set of visual attributes corresponding to the first player in the first track comprises:

generating a plurality of player patches corresponding to the first player based on body-pose information of the first player and an appearance of the first player.

3. The method of claim 2 , wherein learning, by the computing system, the second set of visual attributes corresponding to the second player in the second track comprises:

generating a second plurality of player patches corresponding to the second player based on second body-pose information of the second player and a second appearance of the second player.

4. The method of claim 3 , wherein determining, by the computing system, that the first player and the second player are the same player by measuring the similarity between the first set of visual attributes and the second set of visual attributes comprises:

comparing the plurality of player patches corresponding to the first player to the second plurality of player patches corresponding to the second player.

5. The method of claim 3 , wherein determining, by the computing system, that the first player and the second player are the same player by measuring the similarity between the first set of visual attributes and the second set of visual attributes comprises:

measuring a further similarity between the plurality of player patches and the second plurality of player patches based on their feature representations using a Siamese neural network.

6. The method of claim 1 , further comprising:

based on the inferring, generating, by the computing system, a tracklet of player movement based on the first track and the second track.

7. The method of claim 1 , wherein the first track and the second track do not include any overlapping time.

8. A non-transitory computer readable medium comprising one or more sequences of instructions, which, when executed by a processor, causes a computing system to perform operations comprising:

identifying, by the computing system, broadcast video feed of a game, the broadcast video feed captured by a broadcast camera;

identifying, by the computing system from the broadcast video feed, a first track corresponding to a first player, the first track comprising trajectory information corresponding to the first player up to a first time, wherein after the first time, the first player is not a field of view of the broadcast camera;

identifying, by the computing system from the broadcast video feed, a second track corresponding to a second player, the second track comprising second trajectory information corresponding to the second player up to a second time, the second time after the first time;

learning, by the computing system, a first set of visual attributes corresponding to the first player in the first track;

learning, by the computing system, a second set of visual attributes corresponding to the second player in the second track;

determining, by the computing system, that the first player and the second player are the same player by measuring a similarity between the first set of visual attributes and the second set of visual attributes; and

inferring, by the computing system, movement of the first player between the first time and the second time, wherein the first player is not in the field of view of the broadcast camera between the first time and the second time.

9. The non-transitory computer readable medium of claim 8 , wherein learning, by the computing system, the first set of visual attributes corresponding to the first player in the first track comprises:

generating a plurality of player patches corresponding to the first player based on body-pose information of the first player and an appearance of the first player.

10. The non-transitory computer readable medium of claim 9 , wherein learning, by the computing system, the second set of visual attributes corresponding to the second player in the second track comprises:

generating a second plurality of player patches corresponding to the second player based on second body-pose information of the second player and a second appearance of the second player.

11. The non-transitory computer readable medium of claim 10 , wherein determining, by the computing system, that the first player and the second player are the same player by measuring the similarity between the first set of visual attributes and the second set of visual attributes comprises:

comparing the plurality of player patches corresponding to the first player to the second plurality of player patches corresponding to the second player.

12. The non-transitory computer readable medium of claim 10 , wherein determining, by the computing system, that the first player and the second player are the same player by measuring the similarity between the first set of visual attributes and the second set of visual attributes comprises:

measuring a further similarity between the plurality of player patches and the second plurality of player patches based on their feature representations using a Siamese neural network.

13. The non-transitory computer readable medium of claim 8 , further comprising:

based on the inferring, generating, by the computing system, a tracklet of player movement based on the first track and the second track.

14. The non-transitory computer readable medium of claim 8 , wherein the first track and the second track do not include any overlapping time.

15. A system, comprising:

a processor; and

a memory having programming instructions stored thereon, which, when executed by the processor, causes the system to perform operations comprising:

identifying broadcast video feed of a game, the broadcast video feed captured by a broadcast camera;

identifying, from the broadcast video feed, a first track corresponding to a first player, the first track comprising trajectory information corresponding to the first player up to a first time, wherein after the first time, the first player is not a field of view of the broadcast camera;

identifying, from the broadcast video feed, a second track corresponding to a second player, the second track comprising second trajectory information corresponding to the second player up to a second time, the second time after the first time;

learning a first set of visual attributes corresponding to the first player in the first track;

learning a second set of visual attributes corresponding to the second player in the second track;

determining that the first player and the second player are the same player by measuring a similarity between the first set of visual attributes and the second set of visual attributes; and

inferring movement of the first player between the first time and the second time, wherein the first player is not in the field of view of the broadcast camera between the first time and the second time.

16. The system of claim 15 , wherein learning the first set of visual attributes corresponding to the first player in the first track comprises:

generating a plurality of player patches corresponding to the first player based on body-pose information of the first player and an appearance of the first player.

17. The system of claim 16 , wherein learning the second set of visual attributes corresponding to the second player in the second track comprises:

generating a second plurality of player patches corresponding to the second player based on second body-pose information of the second player and a second appearance of the second player.

18. The system of claim 17 , wherein determining that the first player and the second player are the same player by measuring the similarity between the first set of visual attributes and the second set of visual attributes comprises:

comparing the plurality of player patches corresponding to the first player to the second plurality of player patches corresponding to the second player.

19. The system of claim 17 , wherein determining that the first player and the second player are the same player by measuring the similarity between the first set of visual attributes and the second set of visual attributes comprises:

measuring a further similarity between the plurality of player patches and the second plurality of player patches based on their feature representations using a Siamese neural network.

20. The system of claim 15 , wherein the operations further comprise:

based on the inferring, generating a tracklet of player movement based on the first track and the second track.

Assignments (2)
SECURITY INTEREST Recorded Apr 14, 2026
From: STATS LLC
To: MORGAN STANLEY SENIOR FUNDING, INC., AS COLLATERAL AGENT
Reel/Frame 075390/0491 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 17, 2023
From: SHA, LONG; GANGULY, SUJOY; WEI, XINYU; LUCEY, PATRICK JOSEPH; CHERUKUMUDI, ADITYA
To: STATS LLC
Reel/Frame 062735/0556 →
Continuity (4)
Continuation 17454952 · Nov 15, 2021
Continuation 16805009 · Feb 28, 2020
Provisional Application 62811889 · Feb 28, 2019
Related Publication 20230206464A1 · Jun 29, 2023