Active Scene Understanding via Online Semantic Reconstruction

Zheng, Lintao; Zhu, Chenyang; Zhang, Jiazhao; Zhao, Hang; Huang, Hui; Niessner, Matthias; Xu, Kai

Active Scene Understanding via Online Semantic Reconstruction

dc.contributor.author	Zheng, Lintao	en_US
dc.contributor.author	Zhu, Chenyang	en_US
dc.contributor.author	Zhang, Jiazhao	en_US
dc.contributor.author	Zhao, Hang	en_US
dc.contributor.author	Huang, Hui	en_US
dc.contributor.author	Niessner, Matthias	en_US
dc.contributor.author	Xu, Kai	en_US
dc.contributor.editor	Lee, Jehee and Theobalt, Christian and Wetzstein, Gordon	en_US
dc.date.accessioned	2019-10-14T05:06:44Z
dc.date.available	2019-10-14T05:06:44Z
dc.date.issued	2019
dc.description.abstract	We propose a novel approach to robot-operated active understanding of unknown indoor scenes, based on online RGBD reconstruction with semantic segmentation. In our method, the exploratory robot scanning is both driven by and targeting at the recognition and segmentation of semantic objects from the scene. Our algorithm is built on top of a volumetric depth fusion framework and performs real-time voxel-based semantic labeling over the online reconstructed volume. The robot is guided by an online estimated discrete viewing score field (VSF) parameterized over the 3D space of 2D location and azimuth rotation. VSF stores for each grid the score of the corresponding view, which measures how much it reduces the uncertainty (entropy) of both geometric reconstruction and semantic labeling. Based on VSF, we select the next best views (NBV) as the target for each time step. We then jointly optimize the traverse path and camera trajectory between two adjacent NBVs, through maximizing the integral viewing score (information gain) along path and trajectory. Through extensive evaluation, we show that our method achieves efficient and accurate online scene parsing during exploratory scanning.	en_US
dc.description.number	7
dc.description.sectionheaders	Geometric Modeling
dc.description.seriesinformation	Computer Graphics Forum
dc.description.volume	38
dc.identifier.doi	10.1111/cgf.13820
dc.identifier.issn	1467-8659
dc.identifier.pages	103-114
dc.identifier.uri	https://doi.org/10.1111/cgf.13820
dc.identifier.uri	https://diglib.eg.org:443/handle/10.1111/cgf13820
dc.publisher	The Eurographics Association and John Wiley & Sons Ltd.	en_US
dc.subject	Computing methodologies
dc.subject	Shape analysis
dc.subject	Computer systems organization
dc.subject	Robotic control
dc.title	Active Scene Understanding via Online Semantic Reconstruction	en_US

Collections

38-Issue 7

Active Scene Understanding via Online Semantic Reconstruction

Files

Collections