Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Research paper Hugging Face

DecoEvo: Co-Evolving Solvers and Rubric Generators Without Gold Standards

AI By Crimson AI Hugging Face Papers 30 July 2026 · 00:00 14 views
Share: X Telegram

Hugging Face researchers propose DecoEvo, a framework that co-evolves a solver skill and a rubric-generator skill in text space using decoupled objectives, achieving 2.8–5.0% relative gains over SkillOpt across five benchmarks.

DecoEvo: Co-Evolving Solvers and Rubric Generators Without Gold Standards

Key points

Text-space optimization adapts large language models (LLMs) by editing external natural-language artifacts rather than model weights, keeping the process inspectable and treating the model as a black box. However, most existing methods keep the evaluation fixed, which becomes a bottleneck on open-ended tasks: once the solver improves on the criteria a rubric measures, omitted dimensions remain invisible to the optimization signal.

Simply evolving the rubric is also unreliable when updates are selected by the current solver's score, because apparent progress can come from making the rubric easier to satisfy. To address this, Hugging Face researchers introduce DecoEvo (Decoupled Co-Evolution), which co-evolves a solver skill and a rubric-generator skill under decoupled objectives without using gold rubrics during optimization.

The solver skill is updated using criterion-level feedback, while the rubric-generator skill is revised through complementary audits of requirement coverage and response discrimination that are independent of aggregate solver score. This separation focuses generator updates on newly exposed solver weaknesses, reducing repeated emphasis on criteria the solver already satisfies.

Under each benchmark's official evaluation, DecoEvo outperforms all compared methods across five benchmarks and three LLM backbones, yielding 2.8–5.0% relative gains over SkillOpt in the five-benchmark average. The framework continually refines both problem-solving strategies and evaluation standards by extracting structured feedback from solution audits and analyzing discrepancies across multiple rollouts.

Source
Hugging Face · Hugging Face Papers
Related news
Research paper
Hugging Face 31 Aug 2026

Hugging Face Unveils StepGuard: Step-Level Guardrails for Safer AI Agents

StepGuard, a new step-level guard model from Hugging Face, audits agent actions before execution, reducing attack success rates by...

1
Research paper
Hugging Face 31 Aug 2026

Hugging Face Researchers Unveil ABot-Recon for Stable Long-Horizon 3D Reconstruction

ABot-Recon, a new streaming 3D reconstruction model from Hugging Face, achieves stable long-horizon performance using only local t...

1
Research paper
Hugging Face 31 Aug 2026

ContextPilot: Teaching Agents Proactive Context Management via Fine-Grained RL

Hugging Face researchers introduce ContextPilot, a framework that enhances long-horizon agent reasoning by expanding context-editi...

1