Towards Instance-level Image-to-Image Translation

Shen, Zhiqiang; Huang, Mingyang; Shi, Jianping; Xue, Xiangyang; Huang, Thomas

Computer Science > Computer Vision and Pattern Recognition

arXiv:1905.01744v1 (cs)

[Submitted on 5 May 2019]

Title:Towards Instance-level Image-to-Image Translation

Authors:Zhiqiang Shen, Mingyang Huang, Jianping Shi, Xiangyang Xue, Thomas Huang

View PDF

Abstract:Unpaired Image-to-image Translation is a new rising and challenging vision problem that aims to learn a mapping between unaligned image pairs in diverse domains. Recent advances in this field like MUNIT and DRIT mainly focus on disentangling content and style/attribute from a given image first, then directly adopting the global style to guide the model to synthesize new domain images. However, this kind of approaches severely incurs contradiction if the target domain images are content-rich with multiple discrepant objects. In this paper, we present a simple yet effective instance-aware image-to-image translation approach (INIT), which employs the fine-grained local (instance) and global styles to the target image spatially. The proposed INIT exhibits three import advantages: (1) the instance-level objective loss can help learn a more accurate reconstruction and incorporate diverse attributes of objects; (2) the styles used for target domain of local/global areas are from corresponding spatial regions in source domain, which intuitively is a more reasonable mapping; (3) the joint training process can benefit both fine and coarse granularity and incorporates instance information to improve the quality of global translation. We also collect a large-scale benchmark for the new instance-level translation task. We observe that our synthetic images can even benefit real-world vision tasks like generic object detection.

Comments:	Accepted to CVPR 2019. Project page: this http URL
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI)
Cite as:	arXiv:1905.01744 [cs.CV]
	(or arXiv:1905.01744v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1905.01744

Submission history

From: Zhiqiang Shen [view email]
[v1] Sun, 5 May 2019 20:16:41 UTC (8,914 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Towards Instance-level Image-to-Image Translation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Towards Instance-level Image-to-Image Translation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators