G

Gemini 3.1 Flash Lite Preview

Multimodalby Google·Model page

Google's lightweight multimodal model with 1M-token context for fast text, image, video, and audio tasks.

Rankings
#42on OpenRouter8.7%
633.4B tokens
View rankings
Max output

Most tokens the model can return in a single response.

66Ktokens
Share:

Design Arena

Design Arena ranks models on real-world front-end and design tasks — websites, UI components, data viz, SVG and more — through head-to-head human votes, scored as an ELO rating.
CategoryELOWin rateRank
Asciiart121050.7%#17
UI component111737.7%#74
3D111438.8%#78
Website111236.5%#83
Code110836.4%#81

Pricing

per 1M tokens
Input$0.25 /1M
Output$1.5 /1M
Cache read$0.025 /1M
Cache write$0.083 /1M
Reasoning$1.5 /1M
Audio$0.5 /1M
Image$2.5e-7
Web search$0.01

Model Description

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...

Author
G
Google
Organization · ✓
google
Details
Downloads
Likes
AccessClosed Source
Context1M tokens
Input price$0.25 /1M
Output price$1.5 /1M
CreatedMar 3, 2026
Updated
View on Hugging Face
Benchmarks
Intelligence25.0
Coding34.7
Agentic6.2
Get the full context.

Sign up to read complete case studies, access detailed metrics, and unlock all use cases.

Gemini 3.1 Flash Lite Preview — AI Model Details | Applied