Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
ImpScore: A Learnable Metric For Quantifying The Implicitness Level of Sentences
2025-01-22
· ICLR 2025 Spotlight ·
anchor
Findings
IC-393
GPT-4-turbo, Llama-3.1-8B-Instruct, and OpenAI Moderation show declining hate speech detection accuracy as sentence implicitness increases, with very low success rates in the highest implicitness ranges