UMO: Scaling Multi-Identity Consistency For Image Customization Via Matching Reward - Takara TLDR

Recent advancements in image customization exhibit a wide range of
application prospects due to stronger customization capabilities. However,
since we humans are more sensitive to faces, a significant challenge remains in
preserving consistent identity while avoiding identity confusion with
multi-reference images, limiting the identity scalability of customization
models. To address this, we present UMO, a Unified Multi-identity Optimization
framework, designed to maintain high-fidelity identity preservation and
alleviate identity confusion with scalability. With “multi-to-multi matching”
paradigm, UMO reformulates multi-identity generation as a global assignment
optimization problem and unleashes multi-identity consistency for existing
image customization methods generally through reinforcement learning on
diffusion models. To facilitate the training of UMO, we develop a scalable
customization dataset with multi-reference images, consisting of both
synthesised and real parts. Additionally, we propose a new metric to measure
identity confusion. Extensive experiments demonstrate that UMO not only
improves identity consistency significantly, but also reduces identity
confusion on several image customization methods, setting a new
state-of-the-art among open-source methods along the dimension of identity
preserving. Code and model: https://github.com/bytedance/UMO

Source link

What's Hot

Will there be a ‘September surge’ in hiring this year? Experts weigh in

Perplexity AI tables $34.5 billion cash bid for Google’s Chrome amid Antitrust pressure

Moveworks Recognized as a Challenger in the 2025 Gartner® Magic Quadrant™ for Artificial Intelligence Applications in IT Service Management

UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward – Takara TLDR

Staying in the Sweet Spot: Responsive Reasoning Evolution via Capability-Adaptive Hint Scaffolding – Takara TLDR

F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions – Takara TLDR

Q-Sched: Pushing the Boundaries of Few-Step Diffusion Models with Quantization-Aware Scheduling – Takara TLDR

Growing Support for Parthenon Marbles’ Return to Greece, More Art News

Leon Black and Leslie Wexner’s Letters to Jeffrey Epstein Released

School of Visual Arts Transfers Ownership to Nonprofit Alumni Society

Cristin Tierney Moves Gallery to Tribeca for 15th Anniversary Exhibition

Will there be a ‘September surge’ in hiring this year? Experts weigh in

Perplexity AI tables $34.5 billion cash bid for Google’s Chrome amid Antitrust pressure

Moveworks Recognized as a Challenger in the 2025 Gartner® Magic Quadrant™ for Artificial Intelligence Applications in IT Service Management

What's Hot

UMO: Scaling Multi-Identity Consistency for Image Customization via Matching Reward – Takara TLDR

Related Posts

Subscribe to Updates