Pictorial recognition of objects employing affine invariance in the frequency domain

被引:33
作者
Ben-Arie, J [1 ]
Wang, ZQ [1 ]
机构
[1] Univ Illinois, Dept EECS, Chicago, IL 60607 USA
基金
美国国家科学基金会;
关键词
affine invariant recognition; model-based segmentation; affine invariant spectral signatures (AISS); multidimensional indexing; Gabor kernels;
D O I
10.1109/34.683774
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
This paper describes an efficient approach to pose invariant pictorial object recognition employing spectral signatures of image patches that correspond to object surfaces which are roughly planar. Based on Singular Value Decomposition (SVD), the affine transform is decomposed into slant, tilt, swing, scale, and 2D translation. Unlike previous log-polar representations which were not invariant to slant (i.e., foreshortening only in one direction), our log-log sampling configuration in the frequency domain yields complete affine invariance. The images are preprocessed by a novel model-based segmentation scheme that detects and segments objects that are affine-similar to members of a model set of basic geometric shapes. The segmented objects are then recognized by their signatures using multidimensional indexing in a pictorial dataset represented in the frequency domain. Experimental results with a dataset of 26 models show 100 percent recognition rates in a wide range of 3D pose parameters and imaging degradations: 0-360 degrees swing and tilt, 0-82 degrees of slant (more than 1:7 foreshortening), more than three octaves in scale change, window-limited translation, high noise levels (0 dB), and significantly reduced resolution (1:5).
引用
收藏
页码:604 / 618
页数:15
相关论文
共 28 条
[1]  
[Anonymous], 1996, P EUROPEAN C COMPUTE
[2]   APPLICATION OF AFFINE-INVARIANT FOURIER DESCRIPTORS TO RECOGNITION OF 3-D OBJECTS [J].
ARBTER, K ;
SNYDER, WE ;
BURKHARDT, H ;
HIRZINGER, G .
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 1990, 12 (07) :640-647
[3]   Eigenfaces vs. Fisherfaces: Recognition using class specific linear projection [J].
Belhumeur, PN ;
Hespanha, JP ;
Kriegman, DJ .
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 1997, 19 (07) :711-720
[4]  
BENARIE J, 1996, P 1996 IEEE INT C SP, V6, P3470
[5]  
BENARIE J, 1996, P ARPA IM UND WORKSH, P1277
[6]  
BENARIE J, 1997, 1997 IEEE COMP SOC C
[7]  
BENARIE J, 1996, P IAPR IEEE INT C PA, V1, P672
[8]  
BIEDERMAN I, 1987, COMPUTATIONAL PROCES
[9]  
BUHMANN J, 1992, NEURAL NETWORKS SIGN, P121
[10]   MULTIDIMENSIONAL INDEXING FOR RECOGNIZING VISUAL SHAPES [J].
CALIFANO, A ;
MOHAN, R .
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 1994, 16 (04) :373-392