Latest stories
Hugging Face Unveils MOSS-VL: Open Vision-Language Models for Real-Time Interaction
MOSS-VL, a new open vision-language model family, prioritizes real-time interaction by using gated cross-attention to perceive vis...
Hugging Face Unveils Large Discovery Model: A Recurrent AI Engine for Open-Ended Scientific Search
A new recurrent architecture from Hugging Face combines generative modeling with Bayesian non-parametric surrogates to guide uncer...
SA-MRPO: A Saturation-Aware Approach to Multi-Reward RL for Language Models
Researchers introduce Saturation Aware Advantage Reweighting (SA-MRPO), a method that standardizes each reward objective independe...
StateM Runtime Boosts GPT-5.6 to 95.3% Accuracy on Terminal-Bench 2.1
A new runtime system called StateM improves long-horizon agent performance without changing model weights, achieving 95.3% raw acc...
NVIDIA Scales Expertise with ChatGPT Work
NVIDIA leverages ChatGPT Work to streamline manual tasks, connect dynamic signals, and scale successful workflows across its globa...
GenRouter: Adaptive Routing Cuts Image-Gen Costs by 95%
Hugging Face researchers introduce GenRouter, a unified routing framework that adaptively assigns prompts to optimal agentic image...
Hugging Face Proposes ACID-Compliant Framework for Reliable LLM Agents
A new research paper from Hugging Face introduces 'agentic transactions,' reinterpreting ACID database guarantees for LLM agents,...
UI-Mate: Open-Weight GUI Agent Sets New Benchmarks with In-Context Demonstrations
Hugging Face researchers introduce UI-Mate, a foundation GUI agent that combines environment-grounded training with in-context dem...
Hugging Face Unveils ClawGym II: Black-Box RL Framework for Agent Harness Optimization
A new black-box reinforcement learning framework from Hugging Face enables stable, scalable optimization of general agents through...
HarnessEval-W: Agentic Framework for Transparent World Model Evaluation
Hugging Face researchers introduce HarnessEval-W, an agentified evaluation pipeline that decomposes world-model assessments into v...
VibeWorlding: RL-Trained Open-Source Agents Outperform Closed-Source in 3D World Building
Hugging Face researchers introduce VibeWorlding, a unified framework that benchmarks and trains multimodal agents to construct 3D...
Moonshot AI Unveils Kimi K2.5: Open-Source Visual Agentic Intelligence with Swarm Capabilities
Moonshot AI introduces Kimi K2.5, a native multimodal open-source model with advanced coding and vision, featuring a self-directed...