Skip to main navigation Skip to search Skip to main content

Multi-Domain Image-to-Image Translation with Cross-Granularity Contrastive Learning

  • Huiyuan Fu
  • , Jin Liu
  • , Ting Yu
  • , Xin Wang
  • , Huadong Ma
  • Beijing University of Posts and Telecommunications

Research output: Contribution to journalArticlepeer-review

5 Scopus citations

Abstract

The objective of multi-domain image-to-image translation is to learn the mapping from a source domain to a target domain in multiple image domains while preserving the content representation of the source domain. Despite the importance and recent efforts, most previous studies disregard the large style discrepancy between images and instances in various domains, or fail to capture instance details and boundaries properly, resulting in poor translation results for rich scenes. To address these problems, we present an effective architecture for multi-domain image-to-image translation that only requires one generator. Specifically, we provide detailed procedures for capturing the features of instances throughout the learning process, as well as learning the relationship between the style of the global image and that of a local instance in the image by enforcing the cross-granularity consistency. In order to capture local details within the content space, we employ a dual contrastive learning strategy that operates at both the instance and patch levels. Extensive studies on different multi-domain image-to-image translation datasets reveal that our proposed method outperforms state-of-the-art approaches.

Original languageEnglish
Article number228
JournalACM Transactions on Multimedia Computing, Communications and Applications
Volume20
Issue number7
DOIs
StatePublished - May 16 2024

Keywords

  • contrastive learning
  • cross-granularity
  • GAN
  • Image-to-image translation
  • multi-domain

Fingerprint

Dive into the research topics of 'Multi-Domain Image-to-Image Translation with Cross-Granularity Contrastive Learning'. Together they form a unique fingerprint.

Cite this