Object-Aware 3D Scene Reconstruction from Single 2D Images of Indoor Scenes

被引:2
|
作者
Wen, Mingyun [1 ]
Cho, Kyungeun [1 ]
机构
[1] Dongguk Univ Seoul, Dept Multimedia Engn, 30,Pildong Ro 1 Gil,Jung Gu, Seoul 04620, South Korea
基金
新加坡国家研究基金会;
关键词
3D mesh reconstruction; 3D scene reconstruction; 3D object detection; holistic 3D scene under-standing; deep learning; object-aware reconstruction;
D O I
10.3390/math11020403
中图分类号
O1 [数学];
学科分类号
0701 ; 070101 ;
摘要
Recent studies have shown that deep learning achieves excellent performance in reconstructing 3D scenes from multiview images or videos. However, these reconstructions do not provide the identities of objects, and object identification is necessary for a scene to be functional in virtual reality or interactive applications. The objects in a scene reconstructed as one mesh are treated as a single object, rather than individual entities that can be interacted with or manipulated. Reconstructing an object-aware 3D scene from a single 2D image is challenging because the image conversion process from a 3D scene to a 2D image is irreversible, and the projection from 3D to 2D reduces a dimension. To alleviate the effects of dimension reduction, we proposed a module to generate depth features that can aid the 3D pose estimation of objects. Additionally, we developed a novel approach to mesh reconstruction that combines two decoders that estimate 3D shapes with different shape representations. By leveraging the principles of multitask learning, our approach demonstrated superior performance in generating complete meshes compared to methods relying solely on implicit representation-based mesh reconstruction networks (e.g., local deep implicit functions), as well as producing more accurate shapes compared to previous approaches for mesh reconstruction from single images (e.g., topology modification networks). The proposed method was evaluated on real-world datasets. The results showed that it could effectively improve the object-aware 3D scene reconstruction performance over existing methods.
引用
收藏
页数:16
相关论文
共 50 条
  • [31] 3D Reconstruction of Face from 2D CT Scan Images
    Kumar, T. Senthil
    Vijai, Anupa
    INTERNATIONAL CONFERENCE ON COMMUNICATION TECHNOLOGY AND SYSTEM DESIGN 2011, 2012, 30 : 970 - 977
  • [32] 3D reconstruction from 2D images with hierarchical continuous simplices
    Yunhao Tan
    Jing Hua
    Ming Dong
    The Visual Computer, 2007, 23 : 905 - 914
  • [33] 3D reconstruction and visualization of microstructure surfaces from 2D images
    Samak, D.
    Fischer, A.
    Rittel, D.
    CIRP ANNALS-MANUFACTURING TECHNOLOGY, 2007, 56 (01) : 149 - 152
  • [34] 3D tumor shape reconstruction from 2D bioluminescence images
    Huang, Junzhou
    Huang, Xiaolei
    Metaxas, Dimitris
    Banerjee, Debarata
    2006 3RD IEEE INTERNATIONAL SYMPOSIUM ON BIOMEDICAL IMAGING: MACRO TO NANO, VOLS 1-3, 2006, : 606 - +
  • [35] 3D grape bunch model reconstruction from 2D images
    Woo, Yan San
    Li, Zhuguang
    Tamura, Shun
    Buayai, Prawit
    Nishizaki, Hiromitsu
    Makino, Koji
    Kamarudin, Latifah Munirah
    Mao, Xiaoyang
    COMPUTERS AND ELECTRONICS IN AGRICULTURE, 2023, 215
  • [36] 3D Coordinate Extraction From Single 2D Indoor Image
    Assuja, Maulana Aziz
    Suwardi, Iping Supriana
    2015 INTERNATIONAL SEMINAR ON INTELLIGENT TECHNOLOGY AND ITS APPLICATIONS (ISITIA), 2015, : 233 - 238
  • [37] 3D reconstruction of 2D ICT images by Matlab
    Wei Dongbo
    Li Dan
    Tang Qibo
    Xu Haijun
    ISTM/2007: 7TH INTERNATIONAL SYMPOSIUM ON TEST AND MEASUREMENT, VOLS 1-7, CONFERENCE PROCEEDINGS, 2007, : 3949 - 3952
  • [38] A unified approach to moving object detection in 2D and 3D scenes
    Irani, M
    Anandan, P
    IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 1998, 20 (06) : 577 - 589
  • [39] A unified approach to moving object detection in 2D and 3D scenes
    Irani, M
    Anandan, P
    IMAGE UNDERSTANDING WORKSHOP, 1996 PROCEEDINGS, VOLS I AND II, 1996, : 707 - 718
  • [40] Learning to Synthesize 3D Indoor Scenes from Monocular Images
    Zhu, Fan
    Liu, Li
    Xie, Jin
    Shen, Fumin
    Shao, Ling
    Fang, Yi
    PROCEEDINGS OF THE 2018 ACM MULTIMEDIA CONFERENCE (MM'18), 2018, : 501 - 509