Skip to main content

Write a PREreview

MAVERICK: Breaking the FPR–FNR Seesaw in LLM Output Verification Through Strategy-Differentiated Multi-Agent Consensus

Posted
Server
Preprints.org
DOI
10.20944/preprints202608.1935.v1

The FPR–FNR seesaw—a mathematical constraint on any single-pipeline verifier, where lowering the false-negative rate necessarily raises false positives—is inherent to single-agent judgment, not merely an engineering limitation. We prove that, under conditional error independence, strategy-differentiated multi-agent consensus transcends this trade-off: when agents with deliberately different evidence sources, thresholds and judgment criteria must all agree before an output is automatically released, the system-level FPR falls as p^k—ambiguous cases are escalated to human review rather than silently misclassified. We instantiate this in(MAVERICK), a three-agent architecture, and validate it on LLM reference hallucination (N=5,094): 93.4% unambiguous verdicts,0% FPR on 1,280 fabricated references, and zero misses, consistent with p³ ≤ 2.4×10⁻⁵. Cross-domain pilots on legal citation and clinical trial registry verification (N=400) transfer with zero silent errors among automated verdicts. The architecture is domain-agnostic: wherever LLM outputs can be independently verified, consensus among differentiated verifiers offers a mathematically grounded alternative to model scaling.

You can write a PREreview of MAVERICK: Breaking the FPR–FNR Seesaw in LLM Output Verification Through Strategy-Differentiated Multi-Agent Consensus. A PREreview is a review of a preprint and can vary from a few sentences to a lengthy report, similar to a journal-organized peer-review report.

Before you start

We will ask you to log in with your ORCID iD. If you don’t have an iD, you can create one.

What is an ORCID iD?

An ORCID iD is a unique identifier that distinguishes you from everyone with the same or similar name.

Start now