Theoria: Rewrite-Acceptability Verification over Informal Reasoning States

By Ben Slivinski · Paper · cs.AI

When should an AI system's answer be trusted? Formal proof assistants offer certainty but cannot reach most of the problem distribution; scalar LLM judges offer coverage but produce opaque scores that cannot be audited after the fact and are subject to the same coherence issues a

Cs.ai

View original

HomeResourceLoading…