S2ST: Image-to-Image Translation in the Seed Space of Latent Diffusion

Greenberg, Or; Kishon, Eran; Lischinski, Dani

Computer Science > Computer Vision and Pattern Recognition

arXiv:2312.00116 (cs)

[Submitted on 30 Nov 2023]

Title:S2ST: Image-to-Image Translation in the Seed Space of Latent Diffusion

Authors:Or Greenberg, Eran Kishon, Dani Lischinski

View PDF

Abstract:Image-to-image translation (I2IT) refers to the process of transforming images from a source domain to a target domain while maintaining a fundamental connection in terms of image content. In the past few years, remarkable advancements in I2IT were achieved by Generative Adversarial Networks (GANs), which nevertheless struggle with translations requiring high precision. Recently, Diffusion Models have established themselves as the engine of choice for image generation. In this paper we introduce S2ST, a novel framework designed to accomplish global I2IT in complex photorealistic images, such as day-to-night or clear-to-rain translations of automotive scenes. S2ST operates within the seed space of a Latent Diffusion Model, thereby leveraging the powerful image priors learned by the latter. We show that S2ST surpasses state-of-the-art GAN-based I2IT methods, as well as diffusion-based approaches, for complex automotive scenes, improving fidelity while respecting the target domain's appearance across a variety of domains. Notably, S2ST obviates the necessity for training domain-specific translation networks.

Comments:	17 pages, 15 figures
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Graphics (cs.GR); Machine Learning (cs.LG)
Cite as:	arXiv:2312.00116 [cs.CV]
	(or arXiv:2312.00116v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2312.00116

Submission history

From: Dani Lischinski [view email]
[v1] Thu, 30 Nov 2023 18:59:49 UTC (34,851 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:S2ST: Image-to-Image Translation in the Seed Space of Latent Diffusion

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:S2ST: Image-to-Image Translation in the Seed Space of Latent Diffusion

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators