A Computational Learning Theory of Active Object Recognition Under Uncertainty

被引：26

作者：

Andreopoulos, Alexander ^{[1
]}

Tsotsos, John K. ^{[2
]}

机构：

[1] IBM Res Almaden, San Jose, CA 95120 USA

[2] York Univ, Dept Comp Sci & Engn, Ctr Vis Res, Toronto, ON M3J 2R7, Canada

来源：

INTERNATIONAL JOURNAL OF COMPUTER VISION | 2013年 / 101卷 / 01期

关键词：

Object recognition; Visual search; Active vision; Attention; Computational complexity of vision; VISUAL-ATTENTION; MODEL; COMPLEXITY; SALIENCY; SEARCH; TASK;

D O I：

10.1007/s11263-012-0551-6

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

We present some theoretical results related to the problem of actively searching a 3D scene to determine the positions of one or more pre-specified objects. We investigate the effects that input noise, occlusion, and the VC-dimensions of the related representation classes have in terms of localizing all objects present in the search region, under finite computational resources and a search cost constraint. We present a number of bounds relating the noise-rate of low level feature detection to the VC-dimension of an object representable by an architecture satisfying the given computational constraints. We prove that under certain conditions, the corresponding classes of object localization and recognition problems are efficiently learnable in the presence of noise and under a purposive learning strategy, as there exists a polynomial upper bound on the minimum number of examples necessary to correctly localize the targets under the given models of uncertainty. We also use these arguments to show that passive approaches to the same problem do not necessarily guarantee that the problem is efficiently learnable. Under this formulation, we prove the existence of a number of emergent relations between the object detection noise-rate, the scene representation length, the object class complexity, and the representation class complexity, which demonstrate that selective attention is not only necessary due to computational complexity constraints, but it is also necessary as a noise-suppression mechanism and as a mechanism for efficient object class learning. These results concretely demonstrate the advantages of active, purposive and attentive approaches for solving complex vision problems.

引用

页码：95 / 142

页数：48

共 66 条

[1]

ALOIMONOS J, 1987, INT J COMPUT VISION, V1, P333

[2]

Andreopoulos A., 2009, P INT C COMP VIS

[3] On Sensor Bias in Experimental Methods for Comparing Interest-Point, Saliency, and Recognition Algorithms [J].

Andreopoulos, Alexander ;

Tsotsos, John K. .

IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2012, 34 (01) :110-126

[4] Active 3D Object Localization Using a Humanoid Robot [J].

Andreopoulos, Alexander ;

Hasler, Stephan ;

Wersing, Heiko ;

Janssen, Herbert ;

Tsotsos, John K. ;

Koerner, Edgar .

IEEE TRANSACTIONS ON ROBOTICS, 2011, 27 (01) :47-64

[5]

Angluin D., 1988, Machine Learning, V2, P343, DOI 10.1023/A:1022873112823

[6]

[Anonymous], P INT C COMP VIS

[7]

[Anonymous], P 5 CAN C COMP ROB V

[8]

[Anonymous], COMPUTATIONAL PERSPE, DOI DOI 10.7551/MITPRESS/9780262015417.001.0001

[9]

[Anonymous], P 6 INT JOINT C ART

[10]

[Anonymous], 2013, Perception and communication

← 1 2 3 4 5 6 7 →