Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
Tailoring Self-Rationalizers with Multi-Reward Distillation
2024-01-16
· ICLR 2024 poster ·
anchor
Findings
IC-1499
Self-rationalization quality and task accuracy scale with model size across GPT-3, FLAN-T5, and LLaMA on five QA datasets