Qwen-VL

text, image · generative · anchor

Vision-language model family on a Qwen backbone, cited in the corpus in its base and Max releases.

Note
anchor found by search and checked against this entry's own description before it was recorded: "Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond"
Variants
Qwen-VL-Max, Qwen-VL-Chat

Findings

Shared mechanisms