Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Research paper Hugging Face

LedgerMind: A Provenance-Constrained Framework to Boost Faithfulness in Multimodal Agents

AI By Crimson AI Hugging Face Papers 1 August 2026 · 00:00 16 views
Share: X Telegram

LedgerMind introduces a provenance-constrained state machine for multimodal agents, using a Structured Evidence Ledger to ensure grounded reasoning and reduce hallucinations, improving both accuracy and trajectory-level faithfulness.

LedgerMind: A Provenance-Constrained Framework to Boost Faithfulness in Multimodal Agents

Key points

Multimodal agents for visual question answering are increasingly built as multi-step trajectories that combine perception, retrieval, and reasoning. However, evaluation typically focuses on final-answer accuracy, which cannot reveal whether a correct answer was derived from grounded evidence, language priors, or accidental error cancellation.

To address this, researchers propose treating a multimodal agent trajectory as a provenance-constrained state machine. Tool outputs are normalized into a Structured Evidence Ledger that serves as the trajectory state. Downstream reasoning and decision claims may only cite active ledger entries, and grounding is checked at the entity and numeric level. Repair is realized as typed state transitions that cannot introduce content without tool-produced provenance.

This design is instantiated as LedgerMind, augmented by a Three-Layer Grounding Protocol, an Adaptive Dual-Path Dispatcher that matches reasoning depth to question complexity, and an Event-Triggered Verification-and-Repair engine with a formal provenance non-amplification guarantee.

LedgerMind targets four recurring failure patterns that final-answer accuracy tends to obscure: unsupported intermediate reasoning, citation-backed entity hallucination (Phantom Grounding), over-reasoning on simple queries, and repair-time amplification. Experiments across multiple multimodal reasoning benchmarks and backbone MLLMs show that LedgerMind improves both answer accuracy and trajectory-level faithfulness.

Source
Hugging Face · Hugging Face Papers
Related news
Research paper
Hugging Face 31 Aug 2026

Hugging Face Unveils StepGuard: Step-Level Guardrails for Safer AI Agents

StepGuard, a new step-level guard model from Hugging Face, audits agent actions before execution, reducing attack success rates by...

1
Research paper
Hugging Face 31 Aug 2026

Hugging Face Researchers Unveil ABot-Recon for Stable Long-Horizon 3D Reconstruction

ABot-Recon, a new streaming 3D reconstruction model from Hugging Face, achieves stable long-horizon performance using only local t...

1
Research paper
Hugging Face 31 Aug 2026

ContextPilot: Teaching Agents Proactive Context Management via Fine-Grained RL

Hugging Face researchers introduce ContextPilot, a framework that enhances long-horizon agent reasoning by expanding context-editi...

1