CA2575211C

Apparatus and method for processing video data

Abstract

An apparatus and methods for processing video data are described. The invention provides a representation of video data that can be used to assess agreement between the data and a fitting model for a particular parameterization of the data. This allows the comparison of different parameterization techniques and the selection of the optimum one for continued video processing of the particular data. The representation can be utilized in intermediate form as pad of a larger process or as a feedback mechanism for processing video data. When utilized in its intermediate form, the invention can be used in processes for storage, enhancement, refinement, feature extraction, compression, coding, and transmission of video data. The invention serves to extract salient information in a robust and efficient manner while addressing the problems typically associated with video data sources

CA2575211C, drawing sheet 1
Sheet 1 of 9

Term

Term ended

Expired 28 July 2025, 1.2 years ago.

  1. Priority
  2. Filed
  3. Granted
  4. Expired
  5. Today

10 claims: 3 independent, 7 dependent

  1. 1
    CA 02575211 2010-08-12 CLAIMS:1. An apparatus for the purpose of generating an encoded form of video signal data from a plurality of video frames, comprising: a means of detecting an object in a video frame sequence;a means of tracking said object through two or more frames of the video frame sequence;a means of identifying corresponding elements of said object between two or more frames;a means of modeling such correspondences to generate modeled correspondences;a means of resampling pel data in said video frames associated with said object, said resampling means utilizing said modeled correspondences;a means of segmenting said pel data associated with said object from other pel data in said video frame sequence;a means of decomposing said segmented object pel data said decomposing means comprising Principal Component Analysis, and said segmentation means comprising temporal integration, and said correspondence modeling means comprising a robust sampling consensus for the solution of an affine motion model, and said correspondence modeling means comprising a sampling population based on finite differences generated from block-based motion estimation between two or more video frames in said sequence, and said object detection and tracking means comprising a Viola/Jones face detection algorithm.
  2. 2
    A digital processor apparatus for generating an encoded form of video signal data from a plurality of video frames, comprising:means for detecting an object in a video frame sequence;means for tracking said object through two or more frames of the video frame sequence;means for identifying corresponding elements of said object between two or more video frames;-19CA 02575211 2010-08-12 modeling means for modeling such correspondences and generating a correspondence model;means for resampling pel data corresponding to the object in said video frames, said resampling means utilizing said correspondence model;segmentation means for segmenting said pel data corresponding to said object from other pel data in said video frame sequence, resulting in segmented object pel data;decomposition means for decomposing said segmented object pel data, said decomposition means applying Principal Component Analysis, and said segmentation means including a temporal integration, and said modeling means (i) analyzing the correspondence model using a robust sampling consensus for the solution of an affine motion model, and (ii) analyzing the corresponding elements using a sampling population based on finite differences generated from block-based motion estimation between two or more video frames in said sequence.
  3. 7
    A method of processing video signal data, the video signal data having a video frame sequence, comprising the digital processing steps of:-20CA 02575211 2010-08-12 detecting an object in a video frame sequence, the video frame sequence being from subject video signal data;tracking said object through two or more frames of the video frame sequence;identifying corresponding elements of said object between two or more video frames, said step of identifying resulting in determined correspondences;modeling such determined correspondences and generating a correspondence model;resampling pel data corresponding to the object in said video frames, said resampling utilizing said correspondence model;segmenting said pel data corresponding to said object from other pel data in said video frame sequence, resulting in segmented object pel data, wherein said segmenting includes temporal integration;and decomposing said segmented object pel data, said decomposing using a Principal Component Analysis, wherein the step of segmenting includes: (i) applying block-based motion estimation to the segmented object pel data in multiple video frames, (ii) determining finite differences between two or more video frames, and (iii) generating an affine motion model from the determined finite differences;and said modeling includes (i) analyzing the correspondence model using a robust sampling consensus for the solution of the affine motion model, and (ii) analyzing the corresponding elements using a sampling population based on the determined finite differences.