Regress Before Construct: Regress Autoencoder for Point Cloud Self-supervised Learning

被引:1
|
作者
Liu, Yang [1 ]
Chen, Chen [2 ]
Wang, Can [3 ,4 ]
King, Xulin [5 ]
Liu, Mengyuan [6 ]
机构
[1] Sichuan Univ, Coll Comp Sci, Chengdu, Peoples R China
[2] Univ Cent Florida, Ctr Res Comp Vis, Orlando, FL USA
[3] Univ Kiel, Dept Comp Sci, Lab Multimedia Informat Proc, Kiel, Germany
[4] Hangzhou Linxrobot Co, Hangzhou, Peoples R China
[5] Hangzhou GOTHEN Technol Co Ltd, Hangzhou, Peoples R China
[6] Peking Univ, Shenzhen Grad Sch, Key Lab Machine Percept, Shenzhen, Peoples R China
基金
中国国家自然科学基金;
关键词
point clouds; masked point modeling; self-supervised learning; pre-training;
D O I
10.1145/3581783.3612106
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Masked Autoencoders (MAE) have demonstrated promising performance in self-supervised learning for both 2D and 3D computer vision. Nevertheless, existing MAE-based methods still have certain drawbacks. Firstly, the functional decoupling between the encoder and decoder is incomplete, which limits the encoder's representation learning ability. Secondly, downstream tasks solely utilize the encoder, failing to fully leverage the knowledge acquired through the encoder-decoder architecture in the pre-text task. In this paper, we propose Point Regress AutoEncoder (Point-RAE), a new scheme for regressive autoencoders for point cloud self-supervised learning. The proposed method decouples functions between the decoder and the encoder by introducing a mask regressor, which predicts the masked patch representation from the visible patch representation encoded by the encoder and the decoder reconstructs the target from the predicted masked patch representation. By doing so, we minimize the impact of decoder updates on the representation space of the encoder. Moreover, we introduce an alignment constraint to ensure that the representations for masked patches, predicted from the encoded representations of visible patches, are aligned with the masked patch presentations computed from the encoder. To make full use of the knowledge learned in the pre-training stage, we design a new finetune mode for the proposed Point-RAE. Extensive experiments demonstrate that our approach is efficient during pre-training and generalizes well on various downstream tasks. Specifically, our pre-trained models achieve a high accuracy of 90.28% on the ScanObjectNN hardest split and 94.1% accuracy on ModelNet40, surpassing all the other self-supervised learning methods. Our code and pretrained model are public available at: https://github.com/liuyyy111/Point-RAE.
引用
收藏
页码:1738 / 1749
页数:12
相关论文
共 50 条
  • [41] GMAEEG: A Self-Supervised Graph Masked Autoencoder for EEG Representation Learning
    Fu, Zanhao
    Zhu, Huaiyu
    Zhao, Yisheng
    Huan, Ruohong
    Zhang, Yi
    Chen, Shuohui
    Pan, Yun
    IEEE Journal of Biomedical and Health Informatics, 2024, 28 (11): : 6486 - 6497
  • [42] SPINet: self-supervised point cloud frame interpolation network
    Jiawen Xu
    Xinyi Le
    Cailian Chen
    Xinping Guan
    Neural Computing and Applications, 2023, 35 : 9951 - 9960
  • [43] Feature Guided Masked Autoencoder for Self-Supervised Learning in Remote Sensing
    Wang, Yi
    Hernandez, Hugo Hernandez
    Albrecht, Conrad M.
    Zhu, Xiao Xiang
    IEEE JOURNAL OF SELECTED TOPICS IN APPLIED EARTH OBSERVATIONS AND REMOTE SENSING, 2025, 18 : 321 - 336
  • [44] A Cross Branch Fusion-Based Contrastive Learning Framework for Point Cloud Self-supervised Learning
    Wu, Chengzhi
    Huang, Qianliang
    Jin, Kun
    Pfrommer, Julius
    Beyerer, Juergen
    2024 INTERNATIONAL CONFERENCE IN 3D VISION, 3DV 2024, 2024, : 528 - 538
  • [45] Masked Autoencoder for Self-Supervised Pre-training on Lidar Point Clouds
    Hess, Georg
    Jaxing, Johan
    Svensson, Elias
    Hagerman, David
    Petersson, Christoffer
    Svensson, Lennart
    2023 IEEE/CVF WINTER CONFERENCE ON APPLICATIONS OF COMPUTER VISION WORKSHOPS (WACVW), 2023, : 350 - 359
  • [46] Spatiotemporal Self-supervised Learning for Point Clouds in the Wild
    Wu, Yanhao
    Zhang, Tong
    Ke, Wei
    Susstrunk, Sabine
    Salzmann, Mathieu
    2023 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR, 2023, : 5251 - 5260
  • [47] Self-Supervised Learning for Domain Adaptation on Point Clouds
    Achituve, Idan
    Maron, Haggai
    Chechik, Gal
    2021 IEEE WINTER CONFERENCE ON APPLICATIONS OF COMPUTER VISION (WACV 2021), 2021, : 123 - 133
  • [48] Masked Discrimination for Self-supervised Learning on Point Clouds
    Liu, Haotian
    Cai, Mu
    Lee, Yong Jae
    COMPUTER VISION - ECCV 2022, PT II, 2022, 13662 : 657 - 675
  • [49] Self-Supervised Boundary Point Prediction Task for Point Cloud Domain Adaptation
    Chen, Jintao
    Zhang, Yan
    Huang, Kun
    Ma, Feifan
    Tan, Zhuangbin
    Xu, Zheyu
    IEEE ROBOTICS AND AUTOMATION LETTERS, 2023, 8 (09) : 5878 - 5885
  • [50] Masked Spatio-Temporal Structure Prediction for Self-supervised Learning on Point Cloud Videos
    Shen, Zhiqiang
    Sheng, Xiaoxiao
    Fan, Hehe
    Wang, Longguang
    Guo, Yulan
    Liu, Qiong
    Wen, Hao
    Zhou, Xi
    2023 IEEE/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV 2023), 2023, : 16534 - 16543