Modelpedia
Work in progress
About
Findings
Models
Concepts
Methods
Datasets
Sources
Light
Dark
NYU-v2 / NYU-D
anchor
Findings
IC-035
Removing the inductive bias of locality from Vision Transformers improves or matches performance on classification and regression tasks.
[eval]
IC-981
ImageBind's indirect alignment through images degrades zero-shot performance on non-visual modalities and prevents emergent cross-modal retrieval
[eval]