Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Research paper Hugging Face

ShadowDancer: A New Method to Control Video World Models with Any Action

AI By Crimson AI Hugging Face Papers 31 July 2026 · 00:00 19 views
Share: X Telegram

Hugging Face researchers introduce ShadowDancer, a method that learns unified dynamics representations from video pairs with resampled appearance, enabling any-action, frame-level control of interactive video world models without labels or fine-tuning.

ShadowDancer: A New Method to Control Video World Models with Any Action

Key points

Researchers from Hugging Face have introduced ShadowDancer, a novel approach to achieving any-action, frame-level control of interactive video world models. The method addresses a key representational challenge: existing interfaces either encode actions loosely, leaving the model to improvise, or require structured signals that are difficult to acquire across diverse dynamics.

ShadowDancer leverages demonstration videos, which specify dynamics frame by frame, but overcomes the limitation that a video shows dynamics only through one particular appearance—a single "shadow" of the underlying dynamics. The approach introduces two key innovations: shadow pairs and cross-shadow prediction.

Shadow pairs are video pairs that replay the same dynamics under independently resampled appearance, constructed at scale by the Shadow Library. Cross-shadow prediction learns actions by predicting one shadow from the other, discarding what the pairing resamples and preserving what it preserves, yielding a unified dynamics representation that drives a block-causal world model.

Experiments demonstrate improved action transfer and long action rollout over strong latent-action and interactive world model baselines across diverse dynamics families, with an average blinded win rate of 86% in rollout comparisons. The method works across first/third-person gameplay, open worlds, human motion, camera, and robot manipulation—without labels, motion estimators, or fine-tuning.

Source
Hugging Face · Hugging Face Papers
Related news
Research paper
Hugging Face 31 Aug 2026

Hugging Face Unveils StepGuard: Step-Level Guardrails for Safer AI Agents

StepGuard, a new step-level guard model from Hugging Face, audits agent actions before execution, reducing attack success rates by...

1
Research paper
Hugging Face 31 Aug 2026

Hugging Face Researchers Unveil ABot-Recon for Stable Long-Horizon 3D Reconstruction

ABot-Recon, a new streaming 3D reconstruction model from Hugging Face, achieves stable long-horizon performance using only local t...

1
Research paper
Hugging Face 31 Aug 2026

ContextPilot: Teaching Agents Proactive Context Management via Fine-Grained RL

Hugging Face researchers introduce ContextPilot, a framework that enhances long-horizon agent reasoning by expanding context-editi...

1