G

Gemini 3.5 Flash

Multimodalby Google·Model page

Google's fast reasoning model with a 1M-token context for text, image, video, audio, and document inputs.

Rankings
#23on OpenRouter6.1%
1642B tokens
View rankings
Max output

Most tokens the model can return in a single response.

66Ktokens
Reasoning
Reasoning-first model

Works through a step-by-step chain of thought before answering.

Share:

Use Cases

1A
How Allspice Uses Pinecone to Achieve 97% Ingredient Matching Accuracy
Allspice · Product Development
20% → 97%Ingredient matching accuracy
2WB
How Woven by Toyota Built a 10x Faster Bug Triage Agent for Autonomous Driving
Woven by Toyota · Research & Development
10xBug triage speed and scale improvement

Design Arena

Design Arena ranks models on real-world front-end and design tasks — websites, UI components, data viz, SVG and more — through head-to-head human votes, scored as an ELO rating.
CategoryELOWin rateRank
Game dev131957.6%#11
UI component130858.2%#14
Asciiart130561.5%#5
3D129757.7%#19
SVG129762.3%#3

Pricing

per 1M tokens
Input$1.5 /1M
Output$9 /1M
Cache read$0.15 /1M
Cache write$0.083 /1M
Reasoning$9 /1M
Audio$3 /1M
Image$0.0000015
Web search$0.01

Model Description

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

Author
G
Google
Organization · ✓
google
Details
Downloads
Likes
AccessClosed Source
Context1M tokens
Input price$1.5 /1M
Output price$9 /1M
Knowledge cutoffJan 1, 2025
CreatedMay 19, 2026
Updated
View on Hugging Face
Benchmarks
Intelligence50.2
Coding70.1
Agentic37.4
Get the full context.

Sign up to read complete case studies, access detailed metrics, and unlock all use cases.

Gemini 3.5 Flash — AI Model Details | Applied