Latest stories
GradCuit: New Method Boosts LLM Reasoning by Optimizing Hidden States at Test Time
Researchers introduce GradCuit, a test-time optimization method that directly adjusts latent states in a Transformer layer, achiev...
Hugging Face Unveils WCM: A World Critic Model to Fix Value Estimation in VLA Reinforcement Learning
Researchers at Hugging Face propose the World Critic Model (WCM), a lightweight LeJEPA-based architecture that jointly predicts fu...
SWE-Touch: New Benchmark Reveals Coding Agents Fail When Users Edit Code Mid-Task
Hugging Face researchers introduce SWE-Touch, a benchmark that injects conflicting user edits into coding tasks, showing that even...
Hugging Face Unveils DiffusionGemma: A Diffusion LLM That Generates 1,500 Tokens per Second
DiffusionGemma, an experimental open-weight model from Hugging Face, uses discrete diffusion to generate text in parallel blocks o...
CADENA: A Stepwise AI Approach to Reverse-Engineering CAD Models
Hugging Face researchers introduce CADENA, a model that reconstructs 3D meshes into parametric CAD programs step-by-step, mimickin...
Hugging Face Researchers Tackle Trajectory Anchoring Bias in Autonomous Driving VLMs with DEFT-RLVR
A new paper from Hugging Face researchers identifies a trajectory anchoring bias in autonomous driving VLMs caused by ground-truth...
InfiniSplat: Surface-Aligned 3D Gaussian Splatting from a Single Image
InfiniSplat introduces a surface-aligned representation for feed-forward single-image 3D Gaussian Splatting, using geometry-guided...
Skill-α: RL-Based Progressive Skill Generation Boosts Agent Performance
Researchers introduce Skill-α, a reinforcement learning method that generates high-quality agent skills through progressive editin...
Hugging Face Researchers Propose DAPD to Fix 'Privilege Illusion' in On-Policy Distillation
A new paper from Hugging Face introduces Dual-Anchored Policy Distillation (DAPD), a framework that mitigates the 'privilege illus...
Hugging Face's SKT: Verified Synthetic Data Boosts Agent Skill Use at Scale
A new Hugging Face research paper introduces SKT, a pipeline that generates verified synthetic training data to improve how langua...
Hugging Face Unveils UEmbed: Unified Sparse and Dense Multimodal Embeddings
Hugging Face introduces UEmbed, a decoder-only multimodal embedding model that generates both sparse and dense representations in...
WorldExam: New Benchmark Puts World Models to the Test for Inherent Reactivity
Hugging Face researchers introduce WorldExam, a hierarchical benchmark that evaluates controllable video generation models beyond...