Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Research paper Hugging Face

WithEveryone: New Framework Generates Group Images with Up to 10 Identities

AI By Crimson AI Hugging Face Papers 21 August 2026 · 00:00 11 views
Share: X Telegram

Hugging Face researchers introduce WithEveryone, a unified framework for identity-preserving group image generation that grounds identities to layout plans, improving face similarity and reducing artifacts.

WithEveryone: New Framework Generates Group Images with Up to 10 Identities

Key points

Identity-preserving image generation becomes increasingly unreliable when a scene must contain many specified people. Beyond retaining each identity, the model must bind every reference to a distinct person and location, while training-time identity losses must establish correspondence among several noisy predicted faces.

To address this, researchers at Hugging Face introduce WithEveryone, a unified framework for generating group images with up to ten reference identities. The system injects each selected identity as an addressed token, predicts a structured identity–layout plan, and renders the plan as a visual condition.

Its key objective, Layout-Grounded ID Loss, uses annotated face regions to supervise the intended identities directly, avoiding unstable embedding-based face matching. Additionally, ID Representation Forcing trains a prediction for each identity before image synthesis.

On an identity-disjoint benchmark, WithEveryone achieves the highest target-context identity similarity, improving face similarity from 0.462 (GPT-Image-2) to 0.499, while reducing copy-paste artifacts from 0.169 to 0.055. It covers 97.3% of requested identities with a duplicate rate of only 2.8%.

These results demonstrate that explicit identity–layout grounding enables identity-preserving generation to scale to larger groups without relying on direct reference-face copying.

MetricWithEveryoneGPT-Image-2
Face similarity0.4990.462
Copy-paste artifacts0.0550.169
Identity coverage97.3%-
Duplicate rate2.8%-
Source
Hugging Face · Hugging Face Papers
Related news
Research paper
Hugging Face 29 Aug 2026

Hugging Face Audit: 110 of 124 AI Evaluations Fail to Support Their Claims

A new commit-bound census of 124 Inspect Evals units reveals that 110 stop before deterministic inference due to missing historica...

4
Research paper
Hugging Face 29 Aug 2026

Aphanta: New Framework Diagnoses When Image Editing Boosts Multimodal Reasoning

Hugging Face researchers introduce Aphanta, a diagnostic framework that evaluates when image-editing intermediates improve multimo...

5
Research paper
Hugging Face 29 Aug 2026

EditaLive! Enables Real-Time Character Video Editing for Live Streaming

Hugging Face researchers introduce EditaLive, a framework for real-time human-centric video editing in live streams, achieving sta...

4