Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
StereoSet
anchor
Findings
IC-1005
Social bias neurons in BERT-base-cased and RoBERTa-base are concentrated in the deepest transformer layers
[eval]
IC-1078
LLaMA models (7B through 65B) exhibit gender bias in language generation, coreference resolution, and sentence likelihood, with stereotypical associations driving predictions
[eval]