Learning 3D Semantic Segmentation with only 2D Image Supervision

Kyle Genova,Thomas Funkhouser,Brian Brewington,Forrester Cole,Brian Shucker,Caroline Pantofaru,Xiaoqi Yin,Avneesh Sud,Abhijit Kundu

doi:10.1109/3dv53792.2021.00046

Abstract

With the recent growth of urban mapping and autonomous driving efforts, there has been an explosion of raw 3D data collected from terrestrial platforms with lidar scanners and color cameras. However, due to high labeling costs, ground-truth 3D semantic segmentation annotations are limited in both quantity and geographic diversity, while also being difficult to transfer across sensors. In contrast, large image collections with ground-truth semantic segmentations are readily available for diverse sets of scenes. In this paper, we investigate how to use only those labeled 2D image collections to supervise training 3D semantic segmentation models. Our approach is to train a 3D model from pseudo-labels derived from 2D semantic image segmentations using multi-view fusion. We address several novel issues with this approach, including how to select trusted pseudo-labels, how to sample 3D scenes with rare object categories, and how to decouple input features from 2D images from pseudo-labels during training. The proposed network architecture, 2D3DNet, achieves significantly better performance (+6.2-11.4 mIoU) than baselines during experiments on a new urban dataset with lidar and images captured in 20 cities across 5 continents.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Learning 3D Semantic Segmentation with only 2D Image Supervision

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Active Learning Based 3D Semantic Labeling From Images and Videos
Mengqi Rong ... Zhanyi Hu
IEEE Transactions on Circuits and Systems for Video Technology | VOL. 32
Mengqi Rong, et. al.Mengqi Rong ... Zhanyi Hu
13 May 2021
IEEE Transactions on Circuits and Systems for Video Technology | VOL. 32

3D Perception for Autonomous Mobile Robots Navigation Using Deep Learning for Safe Zones Detection: A Comparative Study
Felipe Manfio Barbosa ... Fernando Santos Osório
-
Felipe Manfio Barbosa, et. al.Felipe Manfio Barbosa ... Fernando Santos Osório
29 Apr 2021
29 Apr 2021

Virtual Multi-view Fusion for 3D Semantic Segmentation
Abhijit Kundu ... Brian Brewington
-
Abhijit Kundu, et. al.Abhijit Kundu ... Brian Brewington
01 Jan 2020
01 Jan 2020

Development of a 3D Semantic Segmentation Camera based on Mask Regional Convolutional Neural Network
Van Luan Tran ... Huei-Yung Lin
-
Van Luan Tran, et. al.Van Luan Tran ... Huei-Yung Lin
01 Jul 2019
01 Jul 2019

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Learning 3D Semantic Segmentation with only 2D Image Supervision

Abstract

Talk to us

Similar Papers