AI model comparison
Open Cost Calculator
Compare AI models side by side.
Pick up to four models and compare the trade-offs that matter: context, price, API access, weights, inputs, outputs, and use cases.
| Attribute | Llama 3.2 90B Vision Instruct Meta |
|---|---|
| Summary | Meta’s large open vision-language model for understanding images and documents. |
| Best for | Document vision, image understanding, self-hosted multimodal RAG |
| Watch out for | Large hardware requirement; primarily understanding rather than image generation. |
| Family | Llama 3.2 |
| Type | Multimodal LLM |
| Status | Open weights |
| Released | 2024-09-25 |
| Context window | 131.1K |
| Max output | Not listed |
| Inputs | Text, image |
| Outputs | Text |
| Reasoning | General reasoning |
| Tool calling | Host dependent |
| Structured output | Host dependent |
| API available | Via Meta and partners |
| Open weights | Yes |
| Self-hostable | Yes |
| License | Llama 3.2 Community License |
| Price | See provider pricing |