Multi-modal tracking of faces for video communications
- 22 November 2002
- conference paper
- Published by Institute of Electrical and Electronics Engineers (IEEE)
- No. 10636919,p. 640-645
- https://doi.org/10.1109/cvpr.1997.609393
Abstract
Visual processes to detect and track faces for video compression and transmission. The system is based on an architecture in which a supervisor selects and activates visual processes in cyclic manner. Control of visual processes is made possible by a confidence factor which accompanies each observation. Fusion of results into a unified estimation for tracking is made possible by estimating a covariance matrix with each observation. Visual processes for face tracking are described using blink detection, normalised color histogram matching, and cross correlation (SSD and NCC). Ensembles of visual processes are organised into processing states so as to provide robust tracking. Transition between states is determined by events detected by processes. The result of face detection is fed into recursive estimator (Kalman filter). The output from the estimator drives a PD controller for a pan/tilt/zoom camera. The resulting system provides robust and precise tracking which operates continuously at approximately 20 images per second on a 150 megahertz computer workstation.Keywords
This publication has 5 references indexed in Scilit:
- Robot vision system with a correlation chip for real-time tracking, optical flow and depth map generationPublished by Institute of Electrical and Electronics Engineers (IEEE) ,2003
- Integration and control of reactive visual processesPublished by Springer Nature ,1994
- Principles and techniques for sensor data fusionSignal Processing, 1993
- Measurement and integration of 3-D structures by tracking edge linesInternational Journal of Computer Vision, 1992
- Color indexingInternational Journal of Computer Vision, 1991