Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Research paper Hugging Face

Hugging Face Unveils Apodex Discovery: A Framework for Verifiable AI-Driven Scientific Investigation

AI By Crimson AI Hugging Face Papers 17 August 2026 · 00:00 9 views
Share: X Telegram

Apodex Discovery introduces a framework for building and evaluating 'discoverative' AI systems that pursue extended, verifiable investigations on real-world scientific problems, surpassing state-of-the-art in AAV capsid design and improving drug repurposing predictions.

Hugging Face Unveils Apodex Discovery: A Framework for Verifiable AI-Driven Scientific Investigation

Key points

Hugging Face researchers have introduced Apodex Discovery, a new framework designed to push AI beyond conventional benchmarks by enabling verifiable, extended investigations into real-world scientific problems. The framework centers on a "heavy-duty solver" that combines a foundation model with tools, harnesses, and control policies to pursue complex, stateful investigations.

The approach is motivated by the Apollo program's success, which relied on explicit objectives, simulation, and iterative correction rather than raw problem-solving ability. Similarly, Apodex Discovery aims to turn ambitious AI goals into structured, verifiable missions. The framework includes a problem-scouting process that surveyed 561 industries across 16 sectors, assembling 423 high-value problems and selecting 20 for initial release.

A key component is the environment-task-episode abstraction, which provides standardized data, tools, constraints, feedback, and verification for intermediate and final outputs. The evaluation metric, HDS6, assesses six dimensions—Tools, Repair, Alternatives, Coherence, Evidence, and Scope—independently of final task success.

In tests, Apodex surpassed published state-of-the-art results in AAV capsid design by 7% across viability, tropism, structure prediction, and generative design. In drug repurposing, the framework improved mean normalized prediction scores by 2.5 and 7.6 points for GPT-5.5 and GPT-5.6-sol, respectively, compared to closed-book baselines. Controlled ablations confirmed that the fixed TRACES episode interface enables clear attribution of performance differences to specific solver components.

The authors position Apodex Discovery as a shift from generative AI to "discoverative AI," targeting problems without ground-truth answers. More details are available at the project website.

MetricResult
Industries surveyed561
Problems assembled423
Problems selected for initial release20
AAV capsid design improvement+7%
Drug repurposing score improvement (GPT-5.5)+2.5 points
Drug repurposing score improvement (GPT-5.6-sol)+7.6 points
Source
Hugging Face · Hugging Face Papers
Related news
Research paper
Hugging Face 29 Aug 2026

Hugging Face Audit: 110 of 124 AI Evaluations Fail to Support Their Claims

A new commit-bound census of 124 Inspect Evals units reveals that 110 stop before deterministic inference due to missing historica...

4
Research paper
Hugging Face 29 Aug 2026

Aphanta: New Framework Diagnoses When Image Editing Boosts Multimodal Reasoning

Hugging Face researchers introduce Aphanta, a diagnostic framework that evaluates when image-editing intermediates improve multimo...

5
Research paper
Hugging Face 29 Aug 2026

EditaLive! Enables Real-Time Character Video Editing for Live Streaming

Hugging Face researchers introduce EditaLive, a framework for real-time human-centric video editing in live streams, achieving sta...

4