Research Papers
Weights vs. Skills: New Survey Maps the Road to Self-Improving Robots
A comprehensive survey from Hugging Face researchers organizes robot learning around two competing paradigms: frozen-weight polici...
Hugging Face Unveils MASS: A New World Model for Scalable Multiplayer Simulation
Researchers at Hugging Face propose MASS, a world model that decouples shared world state from view rendering, enabling consistent...
ContextMaster: Real-Time Multi-Shot Video Creation with Sparse Context Routing
Hugging Face researchers introduce ContextMaster, a unified model that enables real-time interactive multi-shot video creation by...
CalibForge: Using Solver Behavior to Calibrate Terminal Tasks for Agent Training
Hugging Face researchers introduce CalibForge, a system that calibrates terminal tasks using solver feedback to create a 'learnabl...
EffectLearner: New AI Framework Removes Objects and Their Effects in Video
Hugging Face researchers introduce EffectLearner, a semantic-reasoning framework for video object removal that also eliminates obj...
SmartMage: Dynamic Modality Orchestration for 3D Scene Understanding
Hugging Face researchers introduce SmartMage, a unified multimodal LLM that dynamically selects task-relevant modalities for 3D sc...
DyPES-VLA: A New Cross-Embodiment VLA Model with Shared Dynamics Priors
Researchers propose DyPES-VLA, a cross-embodiment VLA model that learns shared dynamics priors via future prediction and uses an e...
Hugging Face Unveils W2-VLA: Task-Conditioned Future Wrist Modeling Boosts Fine-Grained Robot Manipulation
Hugging Face researchers introduce World-to-Wrist VLA (W2-VLA), a vision-language-action model that predicts future wrist states c...
Hugging Face Adapts NVIDIA's Nemotron Stack for Modern Greek RAG, Launches HERA Benchmark
Researchers present an end-to-end adaptation of NVIDIA's Nemotron retrieval stack for Modern Greek, introducing the HERA benchmark...
Invisible Shortcuts: How Vision Encoders Learn Camera Fingerprints
A new study reveals that vision models exploit invisible metadata traces in pixels, learning camera fingerprints and processing cu...
Continual Learning in Transition: From Parameters to System-Level Adaptation
A new research paper from Hugging Face proposes a tri-axial framework—When, How, and Where—to characterize the evolution of contin...
Hugging Face Unveils MameLoshnLM, First Open-Source Yiddish LLM and Benchmark
Researchers introduce MameLoshnLM, an 8B-parameter model tailored for Yiddish, along with new corpora and benchmarks to counter th...