Research Papers
H2R-Bench: New Benchmark Tests AI's Human-to-Robot Video Generation
Hugging Face researchers introduce H2R-Bench, a benchmark for evaluating video generation models that transform human manipulation...
AutoPrune: LLM-Designed Visual Token Pruning Cuts 94.4% Tokens with 99% Performance
Hugging Face researchers introduce AutoPrune, a training-free framework that uses large language models to automatically design vi...
Full-Bandwidth Transformers: Latent Feedback Boosts Reasoning and Efficiency
A new Hugging Face paper introduces full-bandwidth transformers, which feed the top-layer hidden state back into the model via a g...
HPSE: Hybrid-Policy Self-Editing Boosts Composable Knowledge Editing in LLMs
Researchers propose HPSE, a method that improves unstructured knowledge editing by distilling from hybrid rollouts that insert mis...
Instruction Tuning Alters Confidence and Reduces Rationale Diversity Without Improving Calibration
A new study finds that instruction tuning consistently changes model confidence and reduces cross-rationale diversity, yet does no...
LycheeMemory V2: Segment-Level Memory Consolidation Cuts LLM Agent Costs by 86%
Hugging Face researchers introduce LycheeMemory V2, a long-term memory framework that batches interactions into semantic segments,...
Hugging Face Unveils Maglev: A Recurrent Transformer with Sliding Memory for Efficient Long-Context Modeling
Maglev introduces a recurrent Transformer architecture with fixed-size memory that generalizes sliding-window attention, enabling...
Context-Matched Distillation Aligns Teacher Supervision for Autoregressive Video Generation
A new method, Context-Matched Distillation (CMD), aligns teacher supervision with causal generation context in few-step autoregres...
Gambit: Thought-Level Beam Search Boosts Reasoning Efficiency by Up to 68.5%
Hugging Face researchers introduce Gambit, an inference algorithm that uses thought-level beam search to dynamically allocate comp...
LiveAnimate: Real-Time Long-Form Human Animation with 14B Diffusion Transformer
Hugging Face researchers introduce LiveAnimate, the first system to combine real-time streaming with stable long-form human animat...
UniSwap: First Streaming Framework for Joint Audio-Visual Identity Swap in Talking Videos
Hugging Face researchers introduce UniSwap, a unified streaming audio-visual diffusion transformer that simultaneously transfers a...
Massive Activations in Hybrid Linear Attention LLMs: New Study Reveals Pre-Attention Spikes and Inter-Spike Plateaus
A new study from Hugging Face provides the first systematic analysis of massive activations in hybrid linear-attention LLMs, uncov...