IP Library Granted Patent US 9,609,348
Granted Patent B2
US 9,609,348 · App. 14/472,313 · Granted Mar 28, 2017

Systems and methods for video content analysis

Inventors: Fang Shi (San Diego, CA); Jin Ming (ChengDu, CN); Qi Wu (ChengDu, CN); Fan You (San Diego, CA); Kai Bao (Torrance, CA)
Assignee: INTERSIL AMERICAS LLC
H04N19/52H04N19/115H04N19/124H04N19/164H04N19/176H04N19/198H04N19/51H04N19/61H04N5/145
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,609,348
App. No.
14/472,313
Granted
Mar 28, 2017
Kind
B2
Abstract

Video analytics systems and methods are described that typically comprise a video encoder operable to generate macroblock video analytics metadata (VAMD) from a video frame. Functional modules receive the VAMD and an encoded version of the video frame is configured to generate video analytics information related to the frame using the VAMD and the encoded video frame. The downstream decoder can use the VAMD to obtain a global motion vector related to the frame, detect and track motion of an object within the frame and monitor a line provided or found within the frame. Traversals of the line by a moving object can be detected and counted using information in the VAMD and the line may be part of a polygon that delineates an area to be monitored within the encoded frame. The VAMD can comprise macroblock level and video frame level information.

Claims (48)

1. A method comprising a processor, a memory, a video sensor, a video encoder and a transceiver, the memory including instructions stored thereon which, when executed by the processor, perform a method for generating video analytics, the method comprising:

providing, by the processor, information representative of a sequence of images captured by the video sensor to the video encoder that is adapted to encode the information using macroblock-based video encoding to obtain a plurality of video frames, wherein the video sensor and video encoder are co-located in a first apparatus;

generating, by the processor, pixel domain video analytics metadata (VAMD) that includes video content analysis information for each of a plurality of macroblocks while encoding the information in the video encoder;

generating, by the processor, a global video analytics message applicable to a plurality of images in the sequence of images using the VAMD;

generating, by the processor, a local video analytics message applicable to a first video frame in the plurality of video frames using the VAMD; and

transmitting, by the transceiver, the plurality of video frames through a network to a second apparatus with a package comprising the VAMD and the local video analytics message or the global video analytics message,

wherein the second apparatus includes a video analytics processor configured to process the package transmitted by the first apparatus.

2. The method of claim 1 , wherein:

the video analytics processor in the second apparatus is configured to generate video analytics information related to the plurality of video frames based on the local video analytics message or the global video analytics message.

3. The method of claim 1 , wherein the plurality of video frames is obtained by:

compressing the sequence of images.

4. The method of claim 1 , wherein generating the VAMD comprises:

generating motion vectors for the plurality of macroblocks.

5. The method of claim 4 , wherein the motion vectors have sub-pixel granularity.

6. The method of claim 4 , wherein generating the VAMD comprises:

filtering the motion vectors to obtain one filtered motion vector for each of the plurality of macroblocks; and

providing filtered motion vectors in the VAMD.

7. The method of claim 1 , wherein the video encoder is embedded in a communications device.

8. The method of claim 1 , wherein the video encoder is provided in a device that functions as a camera.

9. The method of claim 1 , wherein:

the global video analytics message includes information related to a background frame, a foreground object segmentation descriptor, a camera parameter, predefined motion alarm regions coordination and index, or a virtual line; and

the local video analytics message includes information related to global motion vectors, motion alarm region alarm status, virtual line counting results, object tracking parameters, or camera moving parameters.

10. A device comprising:

a camera;

a video encoder;

a video analytics engine;

a communication interface; and

a video sensor in the camera configured to capture a sequence of images;

the video encoder configured to:

encode the sequence of images in video frames using macroblock-based video encoding to provide encoded video frames, and

generate video analytics metadata (VAMD) that includes video content analysis information for each of a plurality of macroblocks processed the sequence of images in the video frames;

the video analytics engine configured to process the VAMD, and to generate one or more video analytics messages from results obtained by processing the VAMD; and

the communication interface configured to transmit the encoded video frames to a video decoder of a client device, and to transmit the VAMD and the one or more video analytics messages in a layered package to a video analytics processor in the client device that is configured to generate video analytics information related to the sequence of images based on the VAMD, the one or more video analytics messages, and the encoded video frames.

11. The device of claim 10 , wherein the encoded video frames include compressed video frames.

12. The device of claim 10 , wherein the VAMD comprises motion vectors generated for the plurality of macroblocks.

13. The device of claim 10 ,

wherein the one or more video analytics messages comprises at least one global video analytics message applicable to a plurality of images in the sequence of images or at least one local video analytics message applicable to a first video frame in the encoded video frames.

14. The device of claim 10 , wherein the video sensor is provided in a camera.

15. An apparatus comprising:

a camera;

a video encoder configured to provide encoded video frames representative of images received from the camera using macroblock-based video encoding, wherein the video encoder is further configured to generate video analytics metadata (VAMD) that includes video content analysis information for each of a plurality of macroblocks processed while encoding the images;

a video analytics engine configured to process the VAMD, and to generate a global video analytics message applicable to a plurality of images in the images received from the camera or a local video analytics message applicable to one of the encoded video frames; and

a communication interface adapted to transmit the encoded video frames to a video decoder of a client device, and to transmit the VAMD, global video analytics message or the local video analytics message in a layered package to a video analytics processor in the client device that is configured to generate video analytics information related to the images received from the camera based on the VAMD, the global video analytics message or the local video analytics message, and the encoded video frames.

16. The apparatus of claim 15 , wherein the encoded video frames include compressed video frames.

17. The apparatus of claim 15 , wherein the VAMD comprises motion vectors generated for the plurality of macroblocks.

18. The apparatus of claim 15 , wherein the global video analytics message includes information related to a background frame, a foreground object segmentation descriptor, a camera parameter, predefined motion alarm regions coordination and index, or a virtual line.

19. The apparatus of claim 15 , wherein the local video analytics message includes information related to global motion vectors, motion alarm region alarm status, virtual line counting results, object tracking parameters, or camera moving parameters.

20. The apparatus of claim 15 , wherein results obtained by processing the VAMD include information related to motion indexing, background extraction, object segmentation, motion detection, virtual line detection, object counting, motion tracking, speed estimation, a background model, a motion alarm, virtual line detections, or electronic image stabilization parameters.

Assignments (2)
CHANGE OF NAME Recorded Jan 19, 2016
From: INTERSIL AMERICAS INC.
To: INTERSIL AMERICAS LLC
Reel/Frame 037558/0706 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 24, 2015
From: SHI, FANG; MING, JIN; WU, QI; YOU, FAN; BAO, KAI
To: INTERSIL AMERICAS INC.
Reel/Frame 036400/0605 →
Priority Claims (4)
WO PCT/CN2010/076555 · Sep 2, 2010 · international
WO PCT/CN2010/076564 · Sep 2, 2010 · international
WO PCT/CN2010/076567 · Sep 2, 2010 · international
WO PCT/CN2010/076569 · Sep 2, 2010 · international
Continuity (2)
Continuation 13225269 · Sep 2, 2011
Related Publication 20140369417A1 · Dec 18, 2014