Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
On the self-verification limitations of large language models on reasoning and planning tasks
2025-01-22
· ICLR 2025 Poster ·
anchor
Findings
IC-078
GPT-4's self-verification loop causes performance collapse due to high false negative rates in binary verification
IC-079
GPT-4's free-form critique generation is unreliable, containing hallucinated edges, vertex colors, and precondition states
IC-080
GPT-4's performance is largely insensitive to the content of feedback; simple re-prompting with a sound verifier (sampling) matches or exceeds detailed critique