Language-Driven Image Style Transfer

Fu, Tsu-Jui; Wang, Xin Eric; Wang, William Yang

Computer Science > Computer Vision and Pattern Recognition

arXiv:2106.00178 (cs)

[Submitted on 1 Jun 2021 (v1), last revised 17 Jul 2022 (this version, v3)]

Title:Language-Driven Image Style Transfer

Authors:Tsu-Jui Fu, Xin Eric Wang, William Yang Wang

View PDF

Abstract:Despite having promising results, style transfer, which requires preparing style images in advance, may result in lack of creativity and accessibility. Following human instruction, on the other hand, is the most natural way to perform artistic style transfer that can significantly improve controllability for visual effect applications. We introduce a new task, language-driven artistic style transfer (LDAST), to manipulate the style of a content image, guided by a text. We propose contrastive language visual artist (CLVA) that learns to extract visual semantics from style instructions and accomplish LDAST by the patch-wise style discriminator. The discriminator considers the correlation between language and patches of style images or transferred results to jointly embed style instructions. CLVA further compares contrastive pairs of content images and style instructions to improve the mutual relativeness. The results from the same content image can preserve consistent content structures. Besides, they should present analogous style patterns from style instructions that contain similar visual semantics. The experiments show that our CLVA is effective and achieves superb transferred results on LDAST.

Comments:	ECCV'22
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2106.00178 [cs.CV]
	(or arXiv:2106.00178v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2106.00178

Submission history

From: Tsu-Jui Fu [view email]
[v1] Tue, 1 Jun 2021 01:58:50 UTC (38,151 KB)
[v2] Sun, 20 Mar 2022 19:10:12 UTC (28,729 KB)
[v3] Sun, 17 Jul 2022 22:53:37 UTC (28,729 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Language-Driven Image Style Transfer

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Language-Driven Image Style Transfer

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators