← All models

Llama-3.2-11B-Vision-Instruct

llama

Open multimodal Llama model for image understanding, captioning, and visual QA

Family
llama
Providers
4
Context
128,000
Output
4,096
Knowledge
2023-12
Weights
Open
Input
textimage
Output
text
Release date
2024-09-25
Updated
2024-09-25

Providers

Source: models.dev