Passer à la navigation principale Passer à la recherche Passer au contenu principal

SCAM! Transferring Humans Between Images with Semantic Cross Attention Modulation

Résultats de recherche: Le chapitre dans un livre, un rapport, une anthologie ou une collectionContribution à une conférenceRevue par des pairs

Résumé

A large body of recent work targets semantically conditioned image generation. Most such methods focus on the narrower task of pose transfer and ignore the more challenging task of subject transfer that consists in not only transferring the pose but also the appearance and background. In this work, we introduce SCAM (Semantic Cross Attention Modulation), a system that encodes rich and diverse information in each semantic region of the image (including foreground and background), thus achieving precise generation with emphasis on fine details. This is enabled by the Semantic Attention Transformer Encoder that extracts multiple latent vectors for each semantic region, and the corresponding generator that exploits these multiple latents by using semantic cross attention modulation. It is trained only using a reconstruction setup, while subject transfer is performed at test time. Our analysis shows that our proposed architecture is successful at encoding the diversity of appearance in each semantic region. Extensive experiments on the iDesigner, CelebAMask-HD and ADE20K datasets show that SCAM outperforms competing approaches; moreover, it sets the new state of the art on subject transfer.

langue originaleAnglais
titreComputer Vision – ECCV 2022 - 17th European Conference, Proceedings
rédacteurs en chefShai Avidan, Gabriel Brostow, Moustapha Cissé, Giovanni Maria Farinella, Tal Hassner
EditeurSpringer Science and Business Media Deutschland GmbH
Pages713-729
Nombre de pages17
ISBN (imprimé)9783031197802
Les DOIs
étatPublié - 1 janv. 2022
Evénement17th European Conference on Computer Vision, ECCV 2022 - Tel Aviv, Israël
Durée: 23 oct. 202227 oct. 2022

Série de publications

NomLecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
Volume13674 LNCS
ISSN (imprimé)0302-9743
ISSN (Electronique)1611-3349

Une conférence

Une conférence17th European Conference on Computer Vision, ECCV 2022
Pays/TerritoireIsraël
La villeTel Aviv
période23/10/2227/10/22

Empreinte digitale

Examiner les sujets de recherche de « SCAM! Transferring Humans Between Images with Semantic Cross Attention Modulation ». Ensemble, ils forment une empreinte digitale unique.

Contient cette citation