G

Gemini 2.5 Flash Lite Preview 09-2025

Multimodalby Google·Model page

Google's lightweight multimodal model with a 1M-token context supporting text, image, audio, video, and file inputs.

Max output

Most tokens the model can return in a single response.

66Ktokens
Share:

Design Arena

Design Arena ranks models on real-world front-end and design tasks — websites, UI components, data viz, SVG and more — through head-to-head human votes, scored as an ELO rating.
CategoryELOWin rateRank
Website114248.1%#77
Data viz113345.5%#72
Code113047.0%#76
Game dev111345.9%#76
UI component107541.4%#78

Pricing

per 1M tokens
Input$0.1 /1M
Output$0.4 /1M
Cache read$0.01 /1M
Cache write$0.083 /1M
Reasoning$0.4 /1M
Audio$0.3 /1M
Image$1.0e-7
Web search$0.01

Model Description

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

Author
G
Google
Organization · ✓
google
Details
Downloads
Likes
AccessClosed Source
Context1M tokens
Input price$0.1 /1M
Output price$0.4 /1M
Knowledge cutoffJan 31, 2025
CreatedSep 25, 2025
Updated
View on Hugging Face
Get the full context.

Sign up to read complete case studies, access detailed metrics, and unlock all use cases.

Gemini 2.5 Flash Lite Preview 09-2025 — AI Model Details | Applied