Gemini 3.5 Flash
Gemini 3.5 Flash is Google's reasoning-capable model: $1.50/$9 per 1M input/output tokens, 1.0M token context window. It supports native tool calling, extended reasoning, image/file input. Released 2026-05-19.
Input Price
$1.5
Output Price
$9
Cache Read Price
$0.15
Context Window
1.0M
All metrics
| Metric | Value | Confidence | Verified | Provenance receipt |
|---|---|---|---|---|
| Input Price | $1.5 | Verified · 2+ sources | Jun 10, 2026 | Receipt |
| Output Price | $9 | Verified · 2+ sources | Jun 10, 2026 | Receipt |
| Cache Read Price | $0.15 | Verified · 2+ sources | Jun 10, 2026 | Receipt |
| Context Window | 1.0M | Verified · 2+ sources | Jun 10, 2026 | Receipt |
| Max Output Tokens | 66K | Verified · 2+ sources | Jun 10, 2026 | Receipt |
| Tool Calling | Yes | Verified · 2+ sources | Jun 10, 2026 | Receipt |
| Extended Reasoning | Yes | Verified · 2+ sources | Jun 10, 2026 | Receipt |
Pricing
Pay-as-you-go API
$1.5per 1M input tokens
- Output
- $9 / 1M tokens
- Cache read
- $0.15 / 1M tokens
Data sources
Facts on this page are sourced from the following verified sources.
- APImodels.devLast synced 3d ago
Compare Gemini 3.5 Flash
Related models
#1DeepSeek V4 Flash
DeepSeek
DeepSeek V4 Flash is DeepSeek's reasoning-capable model: $0.14/$0.28 per 1M input/output tokens, 1M token context window.
#9Claude Haiku 4.5 (latest)
Anthropic
Claude Haiku 4.5 (latest) is Anthropic's reasoning-capable model: $1/$5 per 1M input/output tokens, 200K token context window, 39.45% on SWE-Bench Pro.
#16Llama 4 Maverick 17B Instruct
Meta
Llama 4 Maverick 17B Instruct is Meta's general-purpose model: 1M token context window, 5.24% on SWE-Bench Pro.
#17o4-mini
OpenAI
o4-mini is OpenAI's reasoning-capable model: $1.10/$4.40 per 1M input/output tokens, 200K token context window.
Spotted an error? Suggest a correction.