Automation & Agents · 29 Sep 2026 · 15:07 CEST
Getting the Source Right, Not Just the Fact: Source-Aware Verification for MCP Agents

Publisher preview · OZZZER analysis pending editorial review.
Our latest paper, ProvenanceGuard: Source-Aware Factuality Verification for MCP-Based LLM Agents (read it on Hugging Face, or on arXiv in the meantime), targets that gap. The failure mode we care about is one we call cross-source conflation: a claim that is true somewhere in the evidence, but attributed to the wrong source. A source-blind verifier may pass it, because the fact does exist in the pool.
A source-aware verifier should not. Consider a customer support agent that answers, "According to the account record, this plan includes a 30-day refund window." The refund window may be perfectly real, but stated in a policy document, not in the account record the answer points to. Pool the two together and the claim looks supported.
Keep them separate and the attribution is wrong, and in a data-sensitive setting a wrong attribution can be as damaging as a wrong fact. The same pattern shows up in a clinical agent, where a patient-specific medication detail taken from a patient-history tool becomes misleading the moment the answer presents it as a finding from the medical literature.
A…
Excerpt supplied by the publisher.
Source
Hugging Face · 29 Sep 2026 · 15:07 CEST
Open the original at Hugging Face ↗