arxivcs.AIcs.CLcs.LGcs.LOcs.SE2026-07-01
Theoria: Rewrite-Acceptability Verification over Informal Reasoning States
Michael Saldivar, Ben Slivinski
When should an AI system's answer be trusted? Formal proof assistants offer certainty but cannot reach most of the problem distribution; scalar LLM judges offer coverage but produce opaque scores that cannot be audited after the fact and are subject to the same coherence issues a…