MiniGPT-4

text, image · generative · anchor

Vision-language model aligning a frozen visual encoder to a frozen language model with a single projection layer.

Note
anchor found by search and checked against this entry's own description: "MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models", whose abstract describes the single projection layer this entry records

Findings

Shared mechanisms