Latest stories
Hugging Face's AVA-Encoder: Agent-Native Video Representation via Knowledge Graphs
AVA-Encoder learns structured video representations via agentic auto-encoding using knowledge graphs, enabling cinematic video gen...
H2R-Bench: New Benchmark Tests AI's Human-to-Robot Video Generation
Hugging Face researchers introduce H2R-Bench, a benchmark for evaluating video generation models that transform human manipulation...
AutoPrune: LLM-Designed Visual Token Pruning Cuts 94.4% Tokens with 99% Performance
Hugging Face researchers introduce AutoPrune, a training-free framework that uses large language models to automatically design vi...
Full-Bandwidth Transformers: Latent Feedback Boosts Reasoning and Efficiency
A new Hugging Face paper introduces full-bandwidth transformers, which feed the top-layer hidden state back into the model via a g...
Anthropic Explains Claude's Invisible Text Watermarking for EU AI Act Compliance
Anthropic details how future Claude models will embed an invisible watermark in generated text to comply with the EU AI Act, witho...
HPSE: Hybrid-Policy Self-Editing Boosts Composable Knowledge Editing in LLMs
Researchers propose HPSE, a method that improves unstructured knowledge editing by distilling from hybrid rollouts that insert mis...
Instruction Tuning Alters Confidence and Reduces Rationale Diversity Without Improving Calibration
A new study finds that instruction tuning consistently changes model confidence and reduces cross-rationale diversity, yet does no...
LycheeMemory V2: Segment-Level Memory Consolidation Cuts LLM Agent Costs by 86%
Hugging Face researchers introduce LycheeMemory V2, a long-term memory framework that batches interactions into semantic segments,...
Hugging Face Unveils Maglev: A Recurrent Transformer with Sliding Memory for Efficient Long-Context Modeling
Maglev introduces a recurrent Transformer architecture with fixed-size memory that generalizes sliding-window attention, enabling...
Context-Matched Distillation Aligns Teacher Supervision for Autoregressive Video Generation
A new method, Context-Matched Distillation (CMD), aligns teacher supervision with causal generation context in few-step autoregres...
Gambit: Thought-Level Beam Search Boosts Reasoning Efficiency by Up to 68.5%
Hugging Face researchers introduce Gambit, an inference algorithm that uses thought-level beam search to dynamically allocate comp...
LiveAnimate: Real-Time Long-Form Human Animation with 14B Diffusion Transformer
Hugging Face researchers introduce LiveAnimate, the first system to combine real-time streaming with stable long-form human animat...