Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
WikiQA
anchor
Findings
IC-1319
Larger LLaMA and LLaMA2 models show better calibration on phrase-level tasks but not consistently on sentence- and paragraph-level tasks
[eval]
IC-282
GPT-2 XL (1.5B) exhibits lower accuracy but reduced overconfidence (smaller ECE and Brier scores) compared to larger models on the CAT benchmark
[eval]
IC-304
Instruction-tuned LMs become more vulnerable to prompt-injected data extraction as model size increases from 7B to 70B
[eval]
IC-305
Mistral-instruct-7b's susceptibility to prompt-injected data extraction follows a U-shaped curve depending on the position of the adversarial prompt within the context window
[eval]