IP Library Granted Patent US 7,751,589
Granted Patent B2
US 7,751,589 · App. 10/581,333 · Granted Jul 6, 2010

Three-dimensional road map estimation from video sequences by tracking pedestrians

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,751,589
App. No.
10/581,333
Granted
Jul 6, 2010
Kind
B2
Abstract

Estimation of a 3D layout of roads and paths traveled by pedestrians is achieved by observing the pedestrians and estimating road parameters from the pedestrian's size and position in a sequence of video frames. The system includes a foreground object detection unit to analyze video frames of a 3D scene and detect objects and object positions in video frames, an object scale prediction unit to estimate 3D transformation parameters for the objects and to predict heights of the objects based at least in part on the parameters, and a road map detection unit to estimate road boundaries of the 3D scene using the object positions to generate the road map.

Claims (32)

1. A method of analyzing video frames capturing a 3D scene over time to automatically generate a road map of the 3D scene comprising:

detecting positions of objects in the video frames;

estimating 3D transformation parameters for the objects;

predicting heights of the objects based at least in part on the 3D transformation parameters;

removing outliers from the predicted heights of objects to produce a filtered set of objects;

using the filtered set of objects to repeat estimating the 3D transformation parameters and to repeat predicting the heights of the objects;

estimating road boundaries of the 3D scene using a background image and the positions of the objects by filling a uniform color region starting from a foot of a position of an object of the objects and stopping when an edge pixel of the background image is reached;

generating the road map;

removing outlier pixels from the road map; and

estimating a height map for objects moving on a road of the road map.

2. The method of claim 1 , wherein detecting positions of objects comprises applying a foreground object detection process to the video frames.

3. The method of claim 1 , wherein estimating road boundaries comprises applying a region growing process to object positions to find pixels of the video frames belonging to a road surface in the 3D scene.

4. An article comprising: a tangible computer-readable storage medium containing instructions, which when executed, result in analyzing video frames capturing a 3D scene over time to automatically generate a road map of the 3D scene by

detecting positions of objects in the video frames;

estimating 3D transformation parameters for the objects;

predicting heights of the objects based at least in part on the 3D transformation parameters;

removing outliers from the predicted heights of objects to produce a filtered set of objects;

using the filtered set of objects to repeat estimating the 3D transformation parameters and to repeat predicting the heights of the objects;

estimating road boundaries of the 3D scene using a background image and the positions of the objects by filling a uniform color region starting from a foot of a position of an object of the objects and stopping when an edge pixel of the background image is reached;

generating the road map;

removing outlier pixels from the road map; and

estimating a height map for objects moving on a road of the road map.

5. The article of claim 4 , wherein instructions for detecting positions of objects comprises instructions for applying a foreground object detection process to the video frames.

6. The article of claim 4 , wherein instructions for estimating road boundaries comprises instructions for applying a region growing process to object positions to find pixels of the video frames belonging to a road surface in the 3D scene.

7. A system comprising:

a foreground object detection unit to analyze video frames of a 3D scene and detect objects and object positions in the video frames;

an object scale prediction unit to estimate 3D transformation parameters for the objects, to predict heights of the objects based at least in part on the 3D transformation parameters, to remove outliers from the predicted heights of objects to produce a filtered set of objects, and to use the filtered set of objects to repeat estimating the 3D transformation parameters and to repeat predicting the heights of the objects; and

a road map detection unit to generate the road map by estimating road boundaries of the 3D scene using the object positions and a background image by filling a uniform color region starting from a foot of a position of an object of the objects and stopping when an edge pixel of the background image is reached, removing outlier pixels from the road map, and estimating a height map for objects moving on a road of the road map.

8. The system of claim 7 , wherein the road map detection unit estimates road boundaries by applying a region growing process to object positions to find pixels of the video frames belonging to a road surface in the 3D scene.

9. The method of claim 1 , wherein the objects comprise a representation of a human being in the video frames.

10. The article of claim 4 , wherein the objects comprise a representation of a human being in the video frames.

11. The system of claim 7 , wherein the objects comprise a representation of a human being in the video frames.

Assignments (6)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 2, 2024
From: UATC, LLC
To: AURORA OPERATIONS, INC.
Reel/Frame 066973/0513 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 29, 2021
From: UBER TECHNOLOGIES, INC.
To: UATC, LLC
Reel/Frame 057649/0301 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 13, 2018
From: INTEL CORPORATION
To: UBER TECHNOLOGIES, INC.
Reel/Frame 044917/0344 →
CORRECTIVE ASSIGNMENT TO CORRECT THE NAMING OF LAST NAMED INVENTOR (KONSTANTIN RODYUSHKIN) AS LAST NAME WAS USED AS FIRST NAME AND FIRST NAME WAS USED AS LAST NAME PREVIOUSLY RECORDED ON REEL 017987 FRAME 0008. ASSIGNOR(S) HEREBY CONFIRMS THE NAME IS KONSTANTIN RODYUSHKIN. Recorded May 21, 2010
From: BOVYRIN, ALEXANDER; RODYUSHKIN, KONSTANTIN
To: INTEL CORPORATION
Reel/Frame 024423/0465 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 31, 2006
From: BOVYRIN, ALEXANDER; KONSTANTIN, RODYUSHKIN
To: INTEL CORPORATION
Reel/Frame 017987/0008 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 16, 2004
From: EPPT, KLAUS PETER; ORBAN, ALEXANDER
To: ABB PATENT GMBH
Reel/Frame 016315/0724 →