跳到主要导航 跳到搜索 跳到主要内容

Disentangling Deep Network for Reconstructing 3D Object Shapes from Single 2D Images

  • Northwestern Polytechnical University Xian
  • Xidian University

科研成果: 书/报告/会议事项章节会议稿件同行评审

2 引用 (Scopus)

摘要

Recovering 3D shapes of deformable objects from single 2D images is an extremely challenging and ill-posed problem. Most existing approaches are based on structure-from-motion or graph inference, where a 3D shape is solved by fitting 2D keypoints/mask instead of directly using the vital cue in the original 2D image. These methods usually require multiple views of an object instance and rely on accurate labeling, detection, and matching of 2D keypoints/mask across multiple images. To overcome these limitations, we make effort to reconstruct 3D deformable object shapes directly from the given unconstrained 2D images. In training, instead of using multiple images per object instance, our approach relaxes the constraint to use images from the same object category (with one 2D image per object instance). The key is to disentangle the category-specific representation of the 3D shape identity and the instance-specific representation of the 3D shape displacement from the 2D training images. In testing, the 3D shape of an object can be reconstructed from the given image by deforming the 3D shape identity according to the 3D shape displacement. To achieve this goal, we propose a novel convolutional encoder-decoder network—the Disentangling Deep Network (DisDN). To demonstrate the effectiveness of the proposed approach, we implement comprehensive experiments on the challenging PASCAL VOC benchmark and use different 3D shape ground-truth in training and testing to avoiding overfitting. The obtained experimental results show that DisDN outperforms other state-of-the-art and baseline methods.

源语言英语
主期刊名Pattern Recognition and Computer Vision - 4th Chinese Conference, PRCV 2021, Proceedings
编辑Huimin Ma, Liang Wang, Changshui Zhang, Fei Wu, Tieniu Tan, Yaonan Wang, Jianhuang Lai, Yao Zhao
出版商Springer Science and Business Media Deutschland GmbH
153-166
页数14
ISBN(印刷版)9783030880064
DOI
出版状态已出版 - 2021
活动4th Chinese Conference on Pattern Recognition and Computer Vision, PRCV 2021 - Beijing, 中国
期限: 29 10月 20211 11月 2021

出版系列

姓名Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
13020 LNCS
ISSN(印刷版)0302-9743
ISSN(电子版)1611-3349

会议

会议4th Chinese Conference on Pattern Recognition and Computer Vision, PRCV 2021
国家/地区中国
Beijing
时期29/10/211/11/21

学术指纹

探究 'Disentangling Deep Network for Reconstructing 3D Object Shapes from Single 2D Images' 的科研主题。它们共同构成独一无二的学术指纹。

引用此