Gemini 2.5 Flash pricing and benchmarks
Google’s previous-generation Flash model.
Input
$0.30
per 1M tokens
Cached input
$0.03
per 1M tokens
Output
$2.50
per 1M tokens
Capability index
140.5
#16 of 16 tracked
How capable is Gemini 2.5 Flash?
Gemini 2.5 Flash scores 140.5 on the Epoch Capabilities Index (confidence interval 138.5–141.9), ranking #16 of the 16 rated models we track.
| Benchmark | Tests | Score |
|---|---|---|
| GPQA Diamond | Science reasoning | — |
| FrontierMath (Tiers 1–3) | Advanced math | — |
| SimpleQA Verified | Factual accuracy | — |
| ARC-AGI-2 | Abstract reasoning | — |
| DeepSWE | Coding | — |
| APEX-Agents | Agentic work | 1.8% |
Source: Epoch AI (as “Gemini 2.5 Flash (Jun 2025)”), best recorded result, CC BY 4.0. A dash means no published score. Compare all models
What Gemini 2.5 Flash costs in practice
| Workload | Monthly cost |
|---|---|
Support chatbot 100,000 replies a month, ~1,500 input tokens (system prompt + history) and ~400 output tokens each. | $145.00 |
Document Q&A (RAG) 50,000 questions a month with ~8,000 tokens of retrieved context and ~600 output tokens each. | $195.00 |
Coding agent 10,000 agent steps a month, ~40,000 input tokens each (70% read from the prompt cache) and ~2,000 output tokens. | $94.40 |
Pricing details
- Audio input costs $1.00 per 1M tokens; text, image and video cost $0.30.
- Output price includes thinking tokens.
Estimate your Gemini 2.5 Flash bill
- Gemini 2.5 FlashGoogle$195.00/mo$0.0039 / request
Standard-tier list prices. Long-context rates apply automatically where the provider publishes a threshold. Excludes cache-write surcharges, taxes, batch and volume discounts.