Towards weakly supervised semantic segmentation in 3D graph-structured point clouds of wild scenes

Haiyan Wang; Xuejian Rong; Liang Yang; Shuihua Wang; Yingli Tian

Towards weakly supervised semantic segmentation in 3D graph-structured point clouds of wild scenes

Haiyan Wang, Xuejian Rong, Liang Yang, Shuihua Wang, Yingli Tian^*

^*Corresponding author for this work

Research output: Contribution to conference › Paper › peer-review

7 Citations (Scopus)

Abstract

The deficiency of 3D segmentation labels is one of the main obstacles to effective point cloud segmentation, especially for wild scenes with varieties of different objects. To alleviate this issue, we propose a novel graph convolutional deep framework for large-scale semantic scene segmentation in point clouds with solely 2D supervision. Different with numerous preceding multi-view supervised approaches focusing on single object point clouds, we argue that 2D supervision is also capable of providing enough guidance information for training 3D semantic segmentation model of natural scene point clouds while not explicitly capturing their inherent structures, even with only single view per sample. Specifically, a Graph-based Pyramid Feature Network (GPFN) is designed to implicitly infer both global and local features of point sets, and a perspective rendering and semantic fusion module are proposed to provide refined 2D supervision signals for training along with a 2D-3D joint optimization strategy. Extensive experimental results demonstrate the effectiveness of our 2D supervised framework, which achieves comparable results with the state-of-the-art approaches trained with full 3D labels, for semantic point cloud segmentation on the popular S3DIS benchmark.

Original language	English
Publication status	Published - 2020
Externally published	Yes
Event	30th British Machine Vision Conference, BMVC 2019 - Cardiff, United Kingdom Duration: 9 Sept 2019 → 12 Sept 2019

Conference

Conference	30th British Machine Vision Conference, BMVC 2019
Country/Territory	United Kingdom
City	Cardiff
Period	9/09/19 → 12/09/19

Cite this

@conference{973136c610764ef98a60efaa1eca3a5b,

title = "Towards weakly supervised semantic segmentation in 3D graph-structured point clouds of wild scenes",

abstract = "The deficiency of 3D segmentation labels is one of the main obstacles to effective point cloud segmentation, especially for wild scenes with varieties of different objects. To alleviate this issue, we propose a novel graph convolutional deep framework for large-scale semantic scene segmentation in point clouds with solely 2D supervision. Different with numerous preceding multi-view supervised approaches focusing on single object point clouds, we argue that 2D supervision is also capable of providing enough guidance information for training 3D semantic segmentation model of natural scene point clouds while not explicitly capturing their inherent structures, even with only single view per sample. Specifically, a Graph-based Pyramid Feature Network (GPFN) is designed to implicitly infer both global and local features of point sets, and a perspective rendering and semantic fusion module are proposed to provide refined 2D supervision signals for training along with a 2D-3D joint optimization strategy. Extensive experimental results demonstrate the effectiveness of our 2D supervised framework, which achieves comparable results with the state-of-the-art approaches trained with full 3D labels, for semantic point cloud segmentation on the popular S3DIS benchmark.",

author = "Haiyan Wang and Xuejian Rong and Liang Yang and Shuihua Wang and Yingli Tian",

note = "Publisher Copyright: {\textcopyright} 2019. The copyright of this document resides with its authors.; 30th British Machine Vision Conference, BMVC 2019 ; Conference date: 09-09-2019 Through 12-09-2019",

year = "2020",

language = "English",

}

TY - CONF

T1 - Towards weakly supervised semantic segmentation in 3D graph-structured point clouds of wild scenes

AU - Wang, Haiyan

AU - Rong, Xuejian

AU - Yang, Liang

AU - Wang, Shuihua

AU - Tian, Yingli

PY - 2020

Y1 - 2020

N2 - The deficiency of 3D segmentation labels is one of the main obstacles to effective point cloud segmentation, especially for wild scenes with varieties of different objects. To alleviate this issue, we propose a novel graph convolutional deep framework for large-scale semantic scene segmentation in point clouds with solely 2D supervision. Different with numerous preceding multi-view supervised approaches focusing on single object point clouds, we argue that 2D supervision is also capable of providing enough guidance information for training 3D semantic segmentation model of natural scene point clouds while not explicitly capturing their inherent structures, even with only single view per sample. Specifically, a Graph-based Pyramid Feature Network (GPFN) is designed to implicitly infer both global and local features of point sets, and a perspective rendering and semantic fusion module are proposed to provide refined 2D supervision signals for training along with a 2D-3D joint optimization strategy. Extensive experimental results demonstrate the effectiveness of our 2D supervised framework, which achieves comparable results with the state-of-the-art approaches trained with full 3D labels, for semantic point cloud segmentation on the popular S3DIS benchmark.

AB - The deficiency of 3D segmentation labels is one of the main obstacles to effective point cloud segmentation, especially for wild scenes with varieties of different objects. To alleviate this issue, we propose a novel graph convolutional deep framework for large-scale semantic scene segmentation in point clouds with solely 2D supervision. Different with numerous preceding multi-view supervised approaches focusing on single object point clouds, we argue that 2D supervision is also capable of providing enough guidance information for training 3D semantic segmentation model of natural scene point clouds while not explicitly capturing their inherent structures, even with only single view per sample. Specifically, a Graph-based Pyramid Feature Network (GPFN) is designed to implicitly infer both global and local features of point sets, and a perspective rendering and semantic fusion module are proposed to provide refined 2D supervision signals for training along with a 2D-3D joint optimization strategy. Extensive experimental results demonstrate the effectiveness of our 2D supervised framework, which achieves comparable results with the state-of-the-art approaches trained with full 3D labels, for semantic point cloud segmentation on the popular S3DIS benchmark.

UR - http://www.scopus.com/inward/record.url?scp=85087328060&partnerID=8YFLogxK

M3 - Paper

AN - SCOPUS:85087328060

T2 - 30th British Machine Vision Conference, BMVC 2019

Y2 - 9 September 2019 through 12 September 2019

ER -

Towards weakly supervised semantic segmentation in 3D graph-structured point clouds of wild scenes

Abstract

Conference

Other files and links

Fingerprint

Cite this