InternLM-XComposer2-VL

text, image · generative · anchor

Vision-language model built on an InternLM backbone for interleaved text and image composition.

Note
anchor found by search and checked against this entry's own description: "InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension", built on InternLM2-7B, which is the backbone the citing paper records

Findings

Shared mechanisms