Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing
2024-01-16
· ICLR 2024 poster ·
anchor
Findings
IC-1007
LLMs cannot reliably self-verify or self-correct their own outputs without external tool feedback
IC-1008
The magnitude of CRITIC's improvement on mathematical program synthesis scales with Llama-2 model size