Gemini 3.5 Flash
Google's fast reasoning model with a 1M-token context for text, image, video, audio, and document inputs.
Use Cases
1
How Allspice Uses Pinecone to Achieve 97% Ingredient Matching Accuracy
Allspice · Product Development
20% → 97%
Ingredient matching accuracy
20% → 97%Ingredient matching accuracy
2
How Woven by Toyota Built a 10x Faster Bug Triage Agent for Autonomous Driving
Woven by Toyota · Research & Development
10x
Bug triage speed and scale improvement
10xBug triage speed and scale improvement
Design Arena
Design Arena ranks models on real-world front-end and design tasks — websites, UI components, data viz, SVG and more — through head-to-head human votes, scored as an ELO rating.| Category | ELO | Win rate | Rank |
|---|---|---|---|
| Game dev | 1319 | 57.6% | #11 |
| UI component | 1308 | 58.2% | #14 |
| Asciiart | 1305 | 61.5% | #5 |
| 3D | 1297 | 57.7% | #19 |
| SVG | 1297 | 62.3% | #3 |
Pricing
per 1M tokensInput$1.5 /1M
Output$9 /1M
Cache read$0.15 /1M
Cache write$0.083 /1M
Reasoning$9 /1M
Audio$3 /1M
Image$0.0000015
Web search$0.01
Model Description
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Get the full context.
Sign up to read complete case studies, access detailed metrics, and unlock all use cases.
Get the full context.
Sign up to read complete case studies, access detailed metrics, and unlock all use cases.