Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Research paper Hugging Face

Hugging Face's BDH-CQ: 150M Model Sets New Cost-Accuracy Frontier on ARC-AGI-1

AI By Crimson AI Hugging Face Papers 11 August 2026 · 00:00 11 views
Share: X Telegram

A new 150M-parameter reasoning model, BDH-CQ, combines in-context learning with recurrent latent reasoning to achieve 29.5% pass@2 on ARC-AGI-1 at a cost of $0.0007 per task, breaking the previous Pareto frontier.

Hugging Face's BDH-CQ: 150M Model Sets New Cost-Accuracy Frontier on ARC-AGI-1

Key points

Hugging Face researchers have introduced BDH-CQ, a novel reasoning model that merges in-context learning with recurrent latent reasoning. Unlike traditional models that verbalize intermediate steps, BDH-CQ performs iterative computation in a high-dimensional latent space, continuously updating its recurrent memory with inputs presented at inference time.

The model was evaluated on the public ARC-AGI-1 benchmark, a challenging test of abstract reasoning. A 150M-parameter configuration achieved 29.5% pass@2 at a computed inference cost of $0.0007 per task. This operating point breaks through the previously reported cost-accuracy Pareto frontier, establishing a new state of the art in benchmark cost efficiency.

The researchers also used controlled ARC-like interventions to analyze what the model learns from demonstrations, how consistently it applies inferred transformations, and which concepts remain difficult. This provides insights into the model's reasoning capabilities and limitations.

BDH-CQ's success suggests that recurrent latent reasoning can be a cost-effective alternative to larger, more verbose models, potentially enabling advanced reasoning on edge devices with limited computational resources.

MetricValue
Parameters150M
ARC-AGI-1 pass@229.5%
Inference cost per task$0.0007
Source
Hugging Face · Hugging Face Papers
Related news
Research paper
Hugging Face 31 Aug 2026

Hugging Face Unveils StepGuard: Step-Level Guardrails for Safer AI Agents

StepGuard, a new step-level guard model from Hugging Face, audits agent actions before execution, reducing attack success rates by...

1
Research paper
Hugging Face 31 Aug 2026

Hugging Face Researchers Unveil ABot-Recon for Stable Long-Horizon 3D Reconstruction

ABot-Recon, a new streaming 3D reconstruction model from Hugging Face, achieves stable long-horizon performance using only local t...

1
Research paper
Hugging Face 31 Aug 2026

ContextPilot: Teaching Agents Proactive Context Management via Fine-Grained RL

Hugging Face researchers introduce ContextPilot, a framework that enhances long-horizon agent reasoning by expanding context-editi...

1