Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Sondos Mohamed

Transfer Learning from Simulated to Real Scenes for Monocular 3D Object Detection

Aug 28, 2024

Sondos Mohamed, Walter Zimmer, Ross Greer, Ahmed Alaaeldin Ghita, Modesto Castrillón-Santana, Mohan Trivedi, Alois Knoll, Salvatore Mario Carta, Mirko Marras

Figure 1 for Transfer Learning from Simulated to Real Scenes for Monocular 3D Object Detection

Figure 2 for Transfer Learning from Simulated to Real Scenes for Monocular 3D Object Detection

Figure 3 for Transfer Learning from Simulated to Real Scenes for Monocular 3D Object Detection

Figure 4 for Transfer Learning from Simulated to Real Scenes for Monocular 3D Object Detection

Abstract:Accurately detecting 3D objects from monocular images in dynamic roadside scenarios remains a challenging problem due to varying camera perspectives and unpredictable scene conditions. This paper introduces a two-stage training strategy to address these challenges. Our approach initially trains a model on the large-scale synthetic dataset, RoadSense3D, which offers a diverse range of scenarios for robust feature learning. Subsequently, we fine-tune the model on a combination of real-world datasets to enhance its adaptability to practical conditions. Experimental results of the Cube R-CNN model on challenging public benchmarks show a remarkable improvement in detection performance, with a mean average precision rising from 0.26 to 12.76 on the TUM Traffic A9 Highway dataset and from 2.09 to 6.60 on the DAIR-V2X-I dataset when performing transfer learning. Code, data, and qualitative video results are available on the project website: https://roadsense3d.github.io.

* 18 pages. Accepted for ECVA European Conference on Computer Vision 2024 (ECCV'24)

Via

Access Paper or Ask Questions