Plain-English Summary
A more deployable Llama 4 model notable for its extremely long context window.
A more deployable Llama 4 model notable for its extremely long context window.
A more deployable Llama 4 model notable for its extremely long context window.
Native multimodal MoE model with 17B active parameters, 16 experts, about 109B total parameters, and up to 10M context.
Very long documents, large codebases, multimodal RAG, self-hosting
The full 10M context may require specialized infrastructure and quality can vary across the window.