Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
RAG
Findings
IC-238
LLMs fail to follow user preferences in zero-shot settings, with accuracy below 10% at 10 turns and near zero at 300 turns
[compared-to]
IC-304
Instruction-tuned LMs become more vulnerable to prompt-injected data extraction as model size increases from 7B to 70B
[context]