Back to AI models
AI model comparison

Compare AI models side by side.

Pick up to four models and compare the trade-offs that matter: context, price, API access, weights, inputs, outputs, and use cases.

Open Cost Calculator
Attribute Llama 3.2 90B Vision Instruct Meta
Summary Meta’s large open vision-language model for understanding images and documents.
Best for Document vision, image understanding, self-hosted multimodal RAG
Watch out for Large hardware requirement; primarily understanding rather than image generation.
Family Llama 3.2
Type Multimodal LLM
Status Open weights
Released 2024-09-25
Context window 131.1K
Max output Not listed
Inputs Text, image
Outputs Text
Reasoning General reasoning
Tool calling Host dependent
Structured output Host dependent
API available Via Meta and partners
Open weights Yes
Self-hostable Yes
License Llama 3.2 Community License
Price See provider pricing