VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks

2025-01-22 · ICLR 2025 Poster · anchor

Findings