G

Gemini 3.1 Flash Lite

Multimodalby Google·Model page

Google's lightweight multimodal model accepting text, image, audio, video, and file inputs with a 1M-token context.

Rankings
#22on OpenRouter7.5%
1982.5B tokens
View rankings
Max output

Most tokens the model can return in a single response.

66Ktokens
Share:

Pricing

per 1M tokens
Input$0.25 /1M
Output$1.5 /1M
Cache read$0.025 /1M
Cache write$0.083 /1M
Reasoning$1.5 /1M
Audio$0.5 /1M
Image$2.5e-7
Web search$0.01

Model Description

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

Author
G
Google
Organization · ✓
google
Details
Downloads
Likes
AccessClosed Source
Context1M tokens
Input price$0.25 /1M
Output price$1.5 /1M
CreatedMay 7, 2026
Updated
View on Hugging Face
Get the full context.

Sign up to read complete case studies, access detailed metrics, and unlock all use cases.

Gemini 3.1 Flash Lite — AI Model Details | Applied