Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Research paper Hugging Face

EditaLive! Enables Real-Time Character Video Editing for Live Streaming

AI By Crimson AI Hugging Face Papers 29 August 2026 · 00:00 4 views
Share: X Telegram

Hugging Face researchers introduce EditaLive, a framework for real-time human-centric video editing in live streams, achieving state-of-the-art performance with faithful facial expression preservation and low-latency inference.

EditaLive! Enables Real-Time Character Video Editing for Live Streaming

Key points

Hugging Face researchers have unveiled EditaLive, a novel framework designed for real-time character video editing in live streaming. Unlike conventional video editing that focuses on scene-level content, EditaLive prioritizes the human subject, addressing the unique challenges of live-streaming environments.

The framework builds on a pretrained image animation model, Wan-Animate, which naturally separates appearance from motion. By repurposing it as the base for instruction-based editing and training on the CharEdit-50K dataset, EditaLive enables reference-frame editing and video reconstruction.

To achieve real-time performance, the researchers adapted the model from offline bidirectional processing to causal streaming generation. They employed a distilled two-step sampling strategy with aligned self-rollout, reducing training-inference discrepancies. Fixed RoPE and align forcing, along with first-frame preserved sparse attention, help mitigate appearance drift and filter redundant historical information.

Extensive experiments demonstrate that EditaLive delivers state-of-the-art editing performance while faithfully preserving facial expressions and maintaining low-latency streaming inference, making it suitable for interactive live-stream applications.

Source
Hugging Face · Hugging Face Papers
Related news
Research paper
Hugging Face 29 Aug 2026

Hugging Face Audit: 110 of 124 AI Evaluations Fail to Support Their Claims

A new commit-bound census of 124 Inspect Evals units reveals that 110 stop before deterministic inference due to missing historica...

3
Research paper
Hugging Face 29 Aug 2026

Aphanta: New Framework Diagnoses When Image Editing Boosts Multimodal Reasoning

Hugging Face researchers introduce Aphanta, a diagnostic framework that evaluates when image-editing intermediates improve multimo...

5
Research paper
Hugging Face 29 Aug 2026

Hugging Face Researchers Unveil Agentic Framework for Consistent Multi-Shot Video Editing

A new agentic framework combining LLMs and VLMs tackles the challenge of editing long multi-shot videos with multiple instructions...

3