Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
Post-hoc Reward Calibration: A Case Study on Length Bias
2025-01-22
· ICLR 2025 Poster ·
anchor
Findings
IC-169
BT-based, DPO-based reward models, and GPT-4 as judge all exhibit significant length bias, with their scores correlating with output length rather than quality