Qwen2-VL

2024-09 · text, image · generative · anchor

Vision-language model handling images at native resolution, cited in the corpus at 2B.

Variants
Qwen2-VL 2B, Qwen2-VL-72B, Qwen2-VL-7B, Qwen2-VL-7B-Instruct

Findings

Shared mechanisms