Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
110 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
44 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Towards 3D Object Detection with 2D Supervision (2211.08287v1)

Published 15 Nov 2022 in cs.CV

Abstract: The great progress of 3D object detectors relies on large-scale data and 3D annotations. The annotation cost for 3D bounding boxes is extremely expensive while the 2D ones are easier and cheaper to collect. In this paper, we introduce a hybrid training framework, enabling us to learn a visual 3D object detector with massive 2D (pseudo) labels, even without 3D annotations. To break through the information bottleneck of 2D clues, we explore a new perspective: Temporal 2D Supervision. We propose a temporal 2D transformation to bridge the 3D predictions with temporal 2D labels. Two steps, including homography wraping and 2D box deduction, are taken to transform the 3D predictions into 2D ones for supervision. Experiments conducted on the nuScenes dataset show strong results (nearly 90% of its fully-supervised performance) with only 25% 3D annotations. We hope our findings can provide new insights for using a large number of 2D annotations for 3D perception.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (6)
  1. Jinrong Yang (27 papers)
  2. Tiancai Wang (48 papers)
  3. Zheng Ge (60 papers)
  4. Weixin Mao (15 papers)
  5. Xiaoping Li (23 papers)
  6. Xiangyu Zhang (328 papers)
Citations (4)

Summary

We haven't generated a summary for this paper yet.