Any-Shot GIN: Generalizing Implicit Networks for Reconstructing Novel Classes

被引：6

作者：

Xian, Yongqin ^{[1
,2
]}

Chibane, Julian ^{[2
,3
]}

Bhatnagar, Bharat Lal ^{[2
,3
]}

Schiele, Bernt ^{[2
]}

Akata, Zeynep ^{[2
,3
,4
]}

Pons-Moll, Gerard ^{[2
,3
]}

机构：

[1] Swiss Fed Inst Technol, Zurich, Switzerland

[2] Max Planck Inst Informat, Saarbrucken, Germany

[3] Univ Tubingen, Tubingen, Germany

[4] Max Planck Inst Intelligent Syst, Stuttgart, Germany

来源：

2022 INTERNATIONAL CONFERENCE ON 3D VISION, 3DV | 2022年

关键词：

D O I：

10.1109/3DV57658.2022.00064

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

We address the task of estimating the 3D shapes of novel shape classes from a single RGB image. Prior works are either limited to reconstructing known training classes or are unable to reconstruct high-quality shapes. To solve those issues, we propose Generalizing Implicit Networks (GIN) which decomposes 3D reconstruction into 1.) front-back depth estimation followed by differentiable depth voxelization, and 2.) implicit shape completion with 3D features. The key insight is that the depth estimation network learns local class-agnostic shape priors, allowing us to generalize to novel classes, while our implicit shape completion network is able to predict accurate shapes with rich details by learning implicit surfaces in 3D voxel space. We conduct extensive experiments on a large-scale benchmark using 55 classes of ShapeNet and real images of Pix3D. We qualitatively and quantitatively show that the proposed GIN significantly outperforms the state of the art on both seen and novel shape classes for single-image 3D reconstruction. We also illustrate that our GIN can be further improved by using only few-shot depth supervision from novel classes.

引用

页码：526 / 535

页数：10

共 55 条

[1] Label-Embedding for Image Classification [J].

Akata, Zeynep ;

Perronnin, Florent ;

Harchaoui, Zaid ;

Schmid, Cordelia .

IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2016, 38 (07) :1425-1438

[2]

Barrow Harry G., 1977, Parametric correspondence and chamfer matching: Two new techniques for image matching

[3]

Bian Wenjing, 2021, BMVC

[4]

Blender Online Community, 2018, Blender-A 3D modelling and rendering package, P6

[5] Geometric Deep Learning Going beyond Euclidean data [J].

Bronstein, Michael M. ;

Bruna, Joan ;

LeCun, Yann ;

Szlam, Arthur ;

Vandergheynst, Pierre .

IEEE SIGNAL PROCESSING MAGAZINE, 2017, 34 (04) :18-42

[6]

Chen W.Y., 2019, INT C LEARNING REPRE

[7] Learning Implicit Fields for Generative Shape Modeling [J].

Chen, Zhiqin ;

Zhang, Hao .

2019 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2019), 2019, :5932-5941

[8]

Cheng Joe, 2024, CRAN

[9]

Chibane J, 2020, ADV NEUR IN, V33

[10] Stereo Radiance Fields (SRF): Learning View Synthesis for Sparse Views of Novel Scenes [J].

Chibane, Julian ;

Bansal, Aayush ;

Lazova, Verica ;

Pons-Moll, Gerard .

2021 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR 2021, 2021, :7907-7916

← 1 2 3 4 5 6 →