← All models

Llama-3.2-11B-Vision-Instruct

llama

Open multimodal Llama model for image understanding, captioning, and visual QA

Context
128,000
Output
4,096
Release date
2024-09-25
Open weights
Yes

Availability from providers