Learning Semantic Segmentation from Synthetic Data: A Geometrically Guided Input-Output Adaptation Approach

Chen, Yuhua; Li, Wen; Chen, Xiaoran; Van Gool, Luc

Computer Science > Computer Vision and Pattern Recognition

arXiv:1812.05040 (cs)

[Submitted on 12 Dec 2018 (v1), last revised 13 Jan 2019 (this version, v2)]

Title:Learning Semantic Segmentation from Synthetic Data: A Geometrically Guided Input-Output Adaptation Approach

Authors:Yuhua Chen, Wen Li, Xiaoran Chen, Luc Van Gool

View PDF

Abstract:Recently, increasing attention has been drawn to training semantic segmentation models using synthetic data and computer-generated annotation. However, domain gap remains a major barrier and prevents models learned from synthetic data from generalizing well to real-world applications. In this work, we take the advantage of additional geometric information from synthetic data, a powerful yet largely neglected cue, to bridge the domain gap. Such geometric information can be generated easily from synthetic data, and is proven to be closely coupled with semantic information. With the geometric information, we propose a model to reduce domain shift on two levels: on the input level, we augment the traditional image translation network with the additional geometric information to translate synthetic images into realistic styles; on the output level, we build a task network which simultaneously performs depth estimation and semantic segmentation on the synthetic data. Meanwhile, we encourage the network to preserve correlation between depth and semantics by adversarial training on the output space. We then validate our method on two pairs of synthetic to real dataset: Virtual KITTI to KITTI, and SYNTHIA to Cityscapes, where we achieve a significant performance gain compared to the non-adapt baseline and methods using only semantic label. This demonstrates the usefulness of geometric information from synthetic data for cross-domain semantic segmentation.

Comments:	v2: fixed some typos
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1812.05040 [cs.CV]
	(or arXiv:1812.05040v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1812.05040

Submission history

From: Yuhua Chen [view email]
[v1] Wed, 12 Dec 2018 17:23:24 UTC (7,790 KB)
[v2] Sun, 13 Jan 2019 23:49:13 UTC (7,791 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Learning Semantic Segmentation from Synthetic Data: A Geometrically Guided Input-Output Adaptation Approach

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Learning Semantic Segmentation from Synthetic Data: A Geometrically Guided Input-Output Adaptation Approach

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators