Research Papers
MMDiff: New Framework Lets Researchers Isolate and Control Features in Multimodal AI Models
Researchers introduce MMDiff, a framework using multimodal sparse autoencoders to identify and control specific features in multim...
Hugging Face Study: Optimal Data Repetition Scales Mildly with LLM Size
A new paper from Hugging Face reveals that under proportional scaling of model size and training tokens, the optimal repetition of...
Hugging Face Paper: Claim-Level Verification Boosts Reasoning Efficiency
A new training-free method, Claim-Level Reliability Assessment (CLR), improves LLM reasoning accuracy by verifying critical claims...
Second Thought: Parallel Reasoning During Agent Idle Time Cuts Sequential Decoding by Up to 43%
A new training-free framework from Hugging Face, Second Thought, runs auxiliary reasoning branches in parallel while LLM agents wa...
LLMs Develop Brain-Like Modular Architecture, Study Finds
A new preprint shows that large language models spontaneously develop modular neural architectures mirroring the human brain's spe...
PRM-as-a-Judge 1.5: New Toolkit for Fine-Grained Robot Process Evaluation
Hugging Face researchers introduce PRM-as-a-Judge 1.5, a toolkit that evaluates embodied robotic models beyond binary success rate...
Hugging Face Researchers Unveil LOPD: A New Self-Distillation Method That Learns Its Own Teaching Context
Latent On-Policy Self-Distillation (LOPD) makes the teacher's privileged context learnable end-to-end, outperforming existing meth...
DFM Mimir v1: Open 1B Model Sets Danish SOTA Using Only Permissible Data
Hugging Face unveils Mimir v1, a 1-billion-parameter Hierarchical Reasoning Model trained solely on permissible data, delivering c...
Hugging Face's Marionette: A New World Model That Predicts 3D States Instead of Pixels
Marionette, a novel world model for interactive games, explicitly predicts 3D articulated states, uses a fixed renderer for geomet...
SimpleOPD: A Tokenizer-Agnostic Approach to Distill Long-Context Reasoning into Smaller Models
Researchers introduce SimpleOPD, a method for on-policy distillation from long-context reasoning teachers to short-context student...
Hugging Face Unveils Apodex Discovery: A Framework for Verifiable AI-Driven Scientific Investigation
Apodex Discovery introduces a framework for building and evaluating 'discoverative' AI systems that pursue extended, verifiable in...
Hugging Face Research: Self-Supervised Distillation Boosts Small VLMs Without Privileged Data
A new self-supervised method, S2VOPD, improves small vision-language models by distilling from original images into strongly augme...