Image-to-Image
Diffusers

MagicMakeup: A Region-Controllable Diffusion Transformer for High-Fidelity Makeup-Transfer

This repository contains the model and code for MagicMakeup: A Region-Controllable Diffusion Transformer for High-Fidelity Makeup-Transfer, as presented in the paper:

overview

Abstract

Makeup-transfer applies the reference makeup to the source face while preserving the source identity. Despite advances in full-face editing by diffusion-based methods, strong regional controllability, makeup fidelity, and identity preservation remain challenging. The reasons are (i) pixel-to-attention misalignment that causes spillover into non-target areas and weakens regional control; (ii) unclear transfer/preservation concept separation under two-image conditioning, leading to coupling between makeup attributes and identity; and (iii) the lack of a high-resolution dataset that is identity-consistent and region-labeled for fine-grained supervision. In this paper, we propose MagicMakeup, a diffusion transformer-based framework for region-controllable and high-fidelity makeup transfer, built on spatial constraints and concept disentanglement. To enable precise region-specific editing while preserving identity, we propose Token-Aligned Region Gating, which aligns pixel masks with attention and applies region-specific logit gating. To clarify the concepts of transfer and preservation, we further introduce Cross-Modal Perception Guidance, which aligns text and image features to enhance cross-modal concept perception. We also design a pipeline for the generation of 1024 x 1024 data pairs through region-specific makeup removal and establish a unified benchmark in synthetic and real settings. Extensive quantitative and qualitative experiments show that MagicMakeup improves regional controllability, makeup fidelity, and identity preservation, with strong robustness across styles, races, and poses.

Code and Usage

The official code and model are available at the following GitHub repository: https://github.com/vivoCameraResearch/Magic-Makeup

Citation

@misc{wang2026magicmakeupregioncontrollablediffusiontransformer, title={MagicMakeup: A Region-Controllable Diffusion Transformer for High-Fidelity Makeup-Transfer}, author={Ziyi Wang and Siming Zheng and Yang Yang and Shusong Xu and Hao Zhang and Bo Li and Changqing Zou and Peng-Tao Jiang}, year={2026}, eprint={2607.20924}, archivePrefix={arXiv}, primaryClass={cs.CV}, url={https://arxiv.org/abs/2607.20924}, }

Downloads last month
134
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Anyou/MagicMakeup

Finetuned
(61)
this model

Paper for Anyou/MagicMakeup