Time3D: End-to-End Joint Monocular 3D Object Detection and Tracking for Autonomous Driving

被引:18
|
作者
Li, Peixuan [1 ]
Jin, Jieyu [1 ]
机构
[1] SAIC PP CEM, Shanghai, Peoples R China
关键词
D O I
10.1109/CVPR52688.2022.00386
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
While separately leveraging monocular 3D object detection and 2D multi-object tracking can be straightforwardly applied to sequence images in a frame-by-frame fashion, stand-alone tracker cuts off the transmission of the uncertainty from the 3D detector to tracking while cannot pass tracking error differentials back to the 3D detector. In this work, we propose jointly training 3D detection and 3D tracking from only monocular videos in an end-to-end manner. The key component is a novel spatial-temporal information flow module that aggregates geometric and appearance features to predict robust similarity scores across all objects in current and past frames. Specifically, we leverage the attention mechanism of the transformer, in which self-attention aggregates the spatial information in a specific frame, and cross-attention exploits relation and affinities of all objects in the temporal domain of sequence frames. The affinities are then supervised to estimate the trajectory and guide the flow of information between corresponding 3D objects. In addition, we propose a temporal -consistency loss that explicitly involves 3D target motion modeling into the learning, making the 3D trajectory smooth in the world coordinate system. Time3D achieves 21.4% AMOTA, 13.6% AMOTP on the nuScenes 3D tracking benchmark, surpassing all published competitors, and running at 38 FPS, while Time3D achieves 31.2% mAP, 39.4% NDS on the nuScenes 3D detection benchmark.
引用
收藏
页码:3875 / 3884
页数:10
相关论文
共 50 条
  • [41] End-to-End Multi-View Fusion for 3D Object Detection in LiDAR Point Clouds
    Zhou, Yin
    Sun, Pei
    Zhang, Yu
    Anguelov, Dragomir
    Gao, Jiyang
    Ouyang, Tom
    Guo, James
    Ngiam, Jiquan
    Vasudevan, Vijay
    CONFERENCE ON ROBOT LEARNING, VOL 100, 2019, 100
  • [42] BirdNet plus : End-to-End 3D Object Detection in LiDAR Bird's Eye View
    Barrera, Alejandro
    Guindel, Carlos
    Beltran, Jorge
    Garcia, Fernando
    2020 IEEE 23RD INTERNATIONAL CONFERENCE ON INTELLIGENT TRANSPORTATION SYSTEMS (ITSC), 2020,
  • [43] 3D Object Detection for Autonomous Driving: A Practical Survey
    Ramajo-Ballester, Alvaro
    de la Escalera Hueso, Arturo
    Armingol Moreno, Jose Maria
    PROCEEDINGS OF THE 9TH INTERNATIONAL CONFERENCE ON VEHICLE TECHNOLOGY AND INTELLIGENT TRANSPORT SYSTEMS, VEHITS 2023, 2023, : 64 - 73
  • [44] 3D Object Detection for Autonomous Driving: A Comprehensive Survey
    Jiageng Mao
    Shaoshuai Shi
    Xiaogang Wang
    Hongsheng Li
    International Journal of Computer Vision, 2023, 131 : 1909 - 1963
  • [45] 3D Object Detection for Autonomous Driving: A Comprehensive Survey
    Mao, Jiageng
    Shi, Shaoshuai
    Wang, Xiaogang
    Li, Hongsheng
    INTERNATIONAL JOURNAL OF COMPUTER VISION, 2023, 131 (08) : 1909 - 1963
  • [46] A review of 3D object detection based on autonomous driving
    Wang, Huijuan
    Chen, Xinyue
    Yuan, Quanbo
    Liu, Peng
    VISUAL COMPUTER, 2024, : 1757 - 1775
  • [47] On Offline Evaluation of 3D Object Detection for Autonomous Driving
    Schreier, Tim
    Renz, Katrin
    Geiger, Andreas
    Chitta, Kashyap
    2023 IEEE/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION WORKSHOPS, ICCVW, 2023, : 4086 - 4091
  • [48] 3D object detection algorithms in autonomous driving: A review
    Ren K.-Y.
    Gu M.-Y.
    Yuan Z.-Q.
    Yuan S.
    Kongzhi yu Juece/Control and Decision, 2023, 38 (04): : 865 - 889
  • [49] Joint Detection and Association for End-to-End Multi-object Tracking
    Li, Ye
    Luo, Xiaoyu
    Shi, Junyu
    Wang, Xinzhong
    Yin, Guangqiang
    Wang, Zhiguo
    NEURAL PROCESSING LETTERS, 2023, 55 (09) : 11823 - 11844
  • [50] Joint Detection and Association for End-to-End Multi-object Tracking
    Ye Li
    Xiaoyu Luo
    Junyu Shi
    Xinzhong Wang
    Guangqiang Yin
    Zhiguo Wang
    Neural Processing Letters, 2023, 55 : 11823 - 11844