Interpreting CLIP's Image Representation via Text-Based Decomposition

2024-01-16 · ICLR 2024 oral · anchor

Findings