Back to AI models

Gemini 3.1 Flash-Lite

A cost-efficient Gemini model for high-volume multimodal applications.

Gemini 3.1 Multimodal LLM Stable Open weights: No API: Yes

Plain-English Summary

A cost-efficient Gemini model for high-volume multimodal applications.

Technical Notes

Native multimodal model with 1,048,576-token input, 65,536-token output, reasoning, tools, code execution, and search grounding.

Best For

High-volume extraction, classification, chat, RAG, media understanding

Watch Out For

Lower quality ceiling than Pro or full Flash models.

Model Specs

Maker
Google
Country
Not listed
Family
Gemini 3.1
Type
Multimodal LLM
Release date
Not listed
Date confidence
Not publicly stated
Status
Stable
Context window
1,048,576 tokens
Max output
65,536
Parameters
Not disclosed
Active parameters
Not disclosed
Architecture
Proprietary native-multimodal transformer
Reasoning
Yes
Tool calling
Yes
Structured output
Yes
API available
Yes
Open weights
No
Self-hostable
No
License
Proprietary
Inputs
Text, image, audio, video, PDF
Outputs
Text
Approx. price
See provider pricing
How to access
Provider website or API
Model / API ID
gemini-3.1-flash-lite
Knowledge cutoff
Not publicly disclosed
Fine-tuning
Prompting / RAG
Completeness
86 %
Verified
Jul 18, 2026
Snapshot date
Jul 18, 2026
Source
Official source