inria-00548634, version 1
Multi-View Object Class Detection with a 3D Geometric Model
Jörg Liebelt 1Cordelia Schmid 2, 3
CVPR 2010 - 23rd IEEE Conference on Computer Vision & Pattern Recognition (2010) 1688-1695
Résumé : This paper presents a new approach for multi-view object class detection. Appearance and geometry are treated as separate learning tasks with different training data. Our approach uses a part model which discriminatively learns the object appearance with spatial pyramids from a database of real images, and encodes the 3D geometry of the object class with a generative representation built from a database of synthetic models. The geometric information is linked to the 2D training data and allows to perform an approximate 3D pose estimation for generic object classes. The pose estimation provides an efficient method to evaluate the likelihood of groups of 2D part detections with respect to a full 3D geometry model in order to disambiguate and prune 2D detections and to handle occlusions. In contrast to other methods, neither tedious manual part annotation of training images nor explicit appearance matching between synthetic and real training data is required, which results in high geometric fidelity and in increased flexibility. On the 3D Object Category datasets CAR and BICYCLE, the current state-of-the-art benchmark for 3D object detection, our approach outperforms previously published results for viewpoint estimation.
- 1 : EADS Innovation Works [Munich] (EADS IW)
- EADS IW
- 2 : LEAR (INRIA Grenoble Rhône-Alpes / LJK Laboratoire Jean Kuntzmann)
- INRIA – Laboratoire Jean Kuntzmann – Université Joseph Fourier - Grenoble I – Institut polytechnique de Grenoble (Grenoble INP) – CNRS : UMR5224
- 3 : Laboratoire Jean Kuntzmann (LJK)
- CNRS : UMR5224 – Université Joseph Fourier - Grenoble I – Université Pierre-Mendès-France - Grenoble II – Institut Polytechnique de Grenoble - Grenoble Institute of Technology
- Domaine : Informatique/Vision par ordinateur et reconnaissance de formes
- Mots-clés : geometry – image representation – object detection – pose estimation – solid modelling
- inria-00548634, version 1
- http://hal.inria.fr/inria-00548634
- oai:hal.inria.fr:inria-00548634
- Contributeur : Team Lear
- Déposé pour le compte de :
- Soumis le : Lundi 20 Décembre 2010, 10:22:35
- Dernière modification le : Vendredi 3 Janvier 2014, 21:12:43