Latest stories
VisCo: Hugging Face Researchers Use LLMs as Intrinsic Encoders for Visual Token Compression
VisCo introduces a training-efficient self-compression framework that reuses a pretrained VLM as an intrinsic compressor, achievin...
Spectral Alignment (SPA): A Lightweight Fix for Exposure Bias in Diffusion Models
Researchers propose Spectral Alignment (SPA), a guidance-based method that calibrates the power spectrum of intermediate predictio...
Training-Free Method Solves Revisit Inconsistency in Autoregressive Video Generation
Researchers propose a training-free approach that uses 3D engine correspondences to maintain consistent appearance when autoregres...
ID-V2V: Netflix and Hugging Face Introduce Identity-Preserving Video Restylization
A new research paper from Netflix and Hugging Face, to appear at SIGGRAPH Asia 2026, presents ID-V2V, a video-to-video framework t...
Multi-Head Latent Control: Lightweight Layer Enables Smarter LLM Agent Decisions
Hugging Face researchers introduce Multi-Head Latent Control, a lightweight layer that reads hidden states from frozen LLMs to pro...
O-VAD: Training-Free Agentic Framework Outperforms Frontier VLMs in Industrial Video Anomaly Detection
Researchers introduce O-VAD, a training-free agentic framework that uses object-centric tracking and reasoning to detect anomalies...
Hugging Face Releases LAMAR: A Language-Aware Multilingual Reranker
LAMAR is a new multilingual cross-encoder reranker that balances semantic relevance and language coherence, outperforming existing...
Hugging Face Researchers Introduce Three-Body Scattering Modeling for One-Step Generation
A new generative framework called Three-Body Scattering Modeling (TBSM) achieves state-of-the-art one-step image generation on Ima...
Interactive Training 2: Open-Source Control Plane Enables Auditable Live Model Steering
Hugging Face introduces Interactive Training 2, an open-source control plane that allows humans and automated agents to adjust tra...
DataPrep-Bench: First Unified Benchmark for LLMs as Training Data Preparators
Hugging Face researchers introduce DataPrep-Bench, the first benchmark jointly evaluating LLMs' data construction and quality eval...
SceneActBench: New Benchmark Tests How Well VLMs Act on 3D Scenes
Hugging Face researchers introduce SceneActBench, a benchmark evaluating vision-language model agents on five 3D tasks within a un...
IDEAgent: Multi-Agent Framework Boosts Research Idea Quality and Diversity by 3.89x
Hugging Face researchers introduce IDEAgent, a multi-agent framework that treats research ideation as a Quality-Diversity search,...