Multi-modal tracking of faces for video communications

被引:66
作者
Crowley, JL
Berard, F
机构
来源
1997 IEEE COMPUTER SOCIETY CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, PROCEEDINGS | 1997年
关键词
D O I
10.1109/CVPR.1997.609393
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
This paper describes a system which multiple visual processes to defect and track faces fur video compression and transmission. The system is based on an architecture in which a supervisor selects and activates visual processes in cyclic manner. control of visual processes is made possible by a confidence factor which accompanies each observation. Fusion of results into a unified estimation for tracking is made possible by estimating a covariance matrix with each observation. Visual processes for face tracking are described using blink detection, normalised color histogram matching, and cross correlation (SSD and NCC). Ensembles of visual processes are organised into processing states so as to provide robust tracking. Transition between states is determined by events detected by processes. The result of face detection is fed into recursive estimator (Kalman filter). The output from the estimator drives a PD controller for a pan/tilt/zoom camera. The resulting system provides robust and precise tracking which operates continuously at approximately 20 images per second on a 150 megahertz computer work-station.
引用
收藏
页码:640 / 645
页数:6
相关论文
empty
未找到相关数据