Saltar al contenido principal

Escribe una PREreview

MAVERICK: Breaking the FPR–FNR Seesaw in LLM Output Verification Through Strategy-Differentiated Multi-Agent Consensus

Publicada
Servidor
Preprints.org
DOI
10.20944/preprints202608.1935.v1

The FPR–FNR seesaw—a mathematical constraint on any single-pipeline verifier, where lowering the false-negative rate necessarily raises false positives—is inherent to single-agent judgment, not merely an engineering limitation. We prove that, under conditional error independence, strategy-differentiated multi-agent consensus transcends this trade-off: when agents with deliberately different evidence sources, thresholds and judgment criteria must all agree before an output is automatically released, the system-level FPR falls as p^k—ambiguous cases are escalated to human review rather than silently misclassified. We instantiate this in(MAVERICK), a three-agent architecture, and validate it on LLM reference hallucination (N=5,094): 93.4% unambiguous verdicts,0% FPR on 1,280 fabricated references, and zero misses, consistent with p³ ≤ 2.4×10⁻⁵. Cross-domain pilots on legal citation and clinical trial registry verification (N=400) transfer with zero silent errors among automated verdicts. The architecture is domain-agnostic: wherever LLM outputs can be independently verified, consensus among differentiated verifiers offers a mathematically grounded alternative to model scaling.

Puedes escribir una PREreview de MAVERICK: Breaking the FPR–FNR Seesaw in LLM Output Verification Through Strategy-Differentiated Multi-Agent Consensus. Una PREreview es una revisión de un preprint y puede variar desde unas pocas oraciones hasta un extenso informe, similar a un informe de revisión por pares organizado por una revista.

Antes de comenzar

We will ask you to log in with your ORCID iD. If you don’t have an iD, you can create one.

What is an ORCID iD?

An ORCID iD is a unique identifier that distinguishes you from everyone with the same or similar name.

Comenzar ahora