Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
LMSYS-Chat-1M: A Large-Scale Real-World LLM Conversation Dataset
2024-01-16
· ICLR 2024 spotlight ·
anchor
Findings
IC-750
Open-source models without safety training are significantly more vulnerable to jailbreak attacks than safety-aligned proprietary models
IC-751
Arena-Hard-200 reveals larger performance gaps between open and proprietary LLMs than MT-Bench
IC-752
GPT-4's win rate over GPT-3.5-turbo is 52% on the top-50 most challenging prompts but only 22% on the bottom-50 easiest prompts