IP Library Granted Patent US 8,655,030
Granted Patent B2
US 8,655,030 · App. 13/467,600 · Granted Feb 18, 2014

Video processing system with face detection and methods for use therewith

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,655,030
App. No.
13/467,600
Granted
Feb 18, 2014
Kind
B2
Abstract

A system for processing a video signal into a processed video signal includes a pattern recognition module for detecting a face in the image sequence, based on coding feedback data, and generating pattern recognition data in response thereto, wherein the pattern recognition data indicates the pattern of interest. A video codec generates the processed video signal and generates the coding feedback data in conjunction with the processing of the image sequence.

Claims (21)

1. A system for processing a video signal into a processed video signal, the video signal including an image sequence, the system comprising:

a pattern recognition module for detecting a face in the image sequence based on coding feedback data and generating pattern recognition data in response thereto, wherein the coding feedback data includes shot transition data that identifies temporal segments in the image sequence corresponding to a plurality of video shots; and

a video codec, coupled to the pattern recognition module, that generates the processed video signal based on the image sequence and by generating the coding feedback data in conjunction with the processing of the image sequence.

2. The system of claim 1 wherein at least one of the plurality of shots includes a plurality of images in the image sequence and wherein the pattern recognition module generates the pattern recognition data based on a temporal recognition performed over the plurality of images.

3. The system of claim 2 wherein temporal recognition tracks a candidate facial region over the plurality of images and detects a facial region based on an identification of facial motion in the candidate facial region over the plurality of images.

4. The system of claim 3 the wherein video codec includes an encoding section that generates the processed video signal by encoding the image sequence, wherein the pattern recognition data includes pattern recognition feedback that indicates the location of the facial region, and wherein the encoder section guides the encoding of the image sequence based on the location of the facial region.

5. The system of claim 3 the wherein the pattern recognition data includes pattern recognition feedback that further indicates the location of at least one of: eyes in the facial region; and a mouth in the facial region.

6. The system of claim 2 wherein temporal recognition tracks a candidate facial region over the plurality of images and extracts three-dimensional features based on different facial perspectives included in the plurality of images.

7. The system of claim 1 wherein the coding feedback data includes at least one image statistic.

8. The system of claim 1 wherein the coding feedback data includes motion vector data.

9. A method for encoding a video signal into a processed video signal, the video signal including an image sequence, the method comprising:

generating encoder feedback data in conjunction with the encoding of the image sequence via an encoder section, wherein the encoder feedback data includes shot transition data that identifies temporal segments in the image sequence corresponding to a plurality of video shots;

detecting a face in the image sequence, based on the encoder feedback data; and

generating pattern recognition data when the face is detected, wherein the pattern recognition data indicates presence of the face.

10. The method of claim 9 wherein at least one of the plurality of shots includes a plurality of images in the image sequence and wherein the pattern recognition data is generated based on a temporal recognition performed over the plurality of images.

11. The method of claim 10 wherein temporal recognition tracks a candidate facial region over the plurality of images and detects a facial region based on an identification of facial motion in the candidate facial region over the plurality of images.

12. The method of claim 11 the wherein the pattern recognition data includes pattern recognition feedback that indicates the location of the facial region, and wherein the encoding of the image sequence is modified based on the location of the facial region.

13. The method of claim 11 the wherein the pattern recognition data includes pattern recognition feedback that further indicates the location of at least one of: eyes in the facial region; and a mouth in the facial region.

14. The method of claim 10 wherein temporal recognition tracks a candidate facial region over the plurality of images and extracts three-dimensional features based on different facial perspectives included in the plurality of images.

15. The method of claim 9 wherein the encoder feedback data includes at least one image statistic.

16. The method of claim 9 wherein the encoder feedback data includes motion vector data.

Assignments (3)
RELEASE OF SECURITY INTEREST Recorded Jul 6, 2023
From: COMERICA BANK
To: VIXS SYSTEMS, INC.
Reel/Frame 064224/0885 →
SECURITY INTEREST Recorded Jul 18, 2016
From: VIXS SYSTEMS INC.
To: COMERICA BANK
Reel/Frame 039380/0479 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 9, 2012
From: LI, YING; ZHAO, XU GANG (WILF)
To: VIXS SYSTEMS, INC.
Reel/Frame 028182/0228 →