Self-Check
Free while signed in. Answers cite the passages they came from.

Explores LLM capacity for self-checking on complex reasoning tasks requiring multi-step and non-linear thinking.
Zero-shot verification: Proposes a zero-shot verification scheme that recognizes errors in its own reasoning without external tools or references.
Weighted voting improvement: Applying self-check scores as weights in majority voting improves QA performance over standard CoT self-consistency.
Math word problems: Demonstrates improved accuracy on math word problems - tasks that benefit most from catching intermediate-step errors.
Self-critique groundwork: An early contribution to the self-critique literature that would mature through 2024 into Constitutional AI-style and debate-style methods.
Get next week’s papers.
The same picks and the same summaries, in your inbox. Free, and 176 issues deep.
Subscribe on Substack