Latest stories
Hugging Face Researchers Unveil TriLayer Dataset and DBL-Diffusion for Layered Video Editing
A new paper from Hugging Face introduces TriLayer, a large-scale video dataset with explicit foreground layers, and DBL-Diffusion,...
Hugging Face Survey Proposes Unified 3D Taxonomy for LLM Memory Architectures
A new survey from Tsinghua, NUS, and Bosch AI, hosted on Hugging Face, introduces a unified framework for understanding memory in...
Voice Memory: A Training-Free Approach to Agentic Speech Recognition
Hugging Face researchers introduce Voice Memory, an inference-only scheme that couples a frozen ASR decoder with a frozen correcto...
New Benchmark OmegaUse-OfficeVal Tests LLM Agents on Cost-Effective Office Tasks
Hugging Face researchers introduce OmegaUse-OfficeVal, a benchmark for evaluating LLM agents on long-horizon office-suite tasks wi...
OVEarth-Bench: New Benchmark Expands Open-Vocabulary Earth Observation Evaluation
Researchers introduce OVEarth-Bench, a unified zero-shot benchmark for open-vocabulary Earth observation that broadens category co...
Shadow Evaluations: A New Test Shows AI Agents Can Engineer but Not Innovate in Research
A new evaluation method, 'shadow evaluations,' reveals that frontier AI agents can complete engineering tasks but fail to make sub...
StatePlay: New AI Model Enforces Game Mechanics for Consistent World Generation
Hugging Face researchers introduce StatePlay, a state-aware game world model that jointly predicts visual content and internal gam...
SpecFirst: A Two-Stage Framework That Boosts From-Scratch Code Synthesis by Up to 21.3%
Hugging Face researchers introduce SpecFirst, a two-stage agent framework that separates behavioral specification elicitation from...
MindForge: Teaching Small Language Models Whole-Life-Cycle Software Engineering via Source-Free Program Synthesis
Hugging Face researchers introduce MindForge, an automated pipeline that converts open-source command-line programs into source-fr...
CoRT: Counterfactual Replay Boosts Token-Level Credit in Rubric-Guided RL
Hugging Face researchers propose CoRT, a token-level credit weighting method for rubric-conditioned GRPO that uses counterfactual...
DistillAlign: A New Approach to Autoregressive Video Distillation Prioritizes Distribution Alignment Over Raw Quality
Researchers propose DistillAlign, a method that coordinates mode covering and mode seeking in autoregressive video distillation, s...
SkillRise: Unified RL Framework Enables LLM Agents to Learn Transferable Skills Across Tasks
Hugging Face researchers introduce SkillRise, a reinforcement learning framework that allows LLM agents to learn and reuse skills...