Claude Haiku 4.5 vs Gemini 3.5 Flash: pricing and benchmarks
Claude Haiku 4.5 is 33% cheaper on input tokens ($1.00 vs $1.50 per 1M). Claude Haiku 4.5 is 44% cheaper on output tokens ($5.00 vs $9.00 per 1M). On our document Q&A workload (50,000 questions a month), Claude Haiku 4.5 costs $550.00 a month versus $870.00 for Gemini 3.5 Flash, 37% less. Gemini 3.5 Flash scores 154.7 on the Epoch Capabilities Index vs 142.4 for Claude Haiku 4.5.
Prices verified September 14, 2026.
Which should you choose?
Gemini 3.5 Flash scores 154.7 on the Epoch Capabilities Index vs 142.4 for Claude Haiku 4.5.
Claude Haiku 4.5 costs $2.00 per 1M tokens (3:1 input-to-output mix) vs $3.375 for Gemini 3.5 Flash, 41% less.
A trade-off: Gemini 3.5 Flash is 12.3 ECI points more capable but costs 1.7x as much as Claude Haiku 4.5.
Pricing and specs
| Claude Haiku 4.5 | Gemini 3.5 Flash | |
|---|---|---|
| Provider | Anthropic | |
| API model ID | claude-haiku-4-5-20251001 | gemini-3.5-flash |
| Input / 1M tokens | $1.00 | $1.50 |
| Cached input / 1M tokens | $0.10 | $0.15 |
| Output / 1M tokens | $5.00 | $9.00 |
| Long-context rates | — | — |
| Context window | 200K tokens | 1.05M tokens |
| Max output | 64K tokens | 66K tokens |
| Batch discount | 50% off | 50% off |
Benchmarks
| Claude Haiku 4.5 | Gemini 3.5 Flash | |
|---|---|---|
| Epoch Capabilities Index | 142.4 | 154.7 |
| GPQA Diamond· Science reasoning | 61.6% | 90.4% |
| FrontierMath (Tiers 1–3)· Advanced math | — | 62.8% |
| SimpleQA Verified· Factual accuracy | 13.2% | 66.2% |
| ARC-AGI-2· Abstract reasoning | 4.0% | 72.1% |
| DeepSWE· Coding | — | 37.4% |
| APEX-Agents· Agentic work | 8.9% | — |
Source: Epoch AI, best recorded result per model (CC BY 4.0). A dash means no published score. See all benchmarks
Monthly cost for real workloads
| Workload | Claude Haiku 4.5 | Gemini 3.5 Flash | Difference |
|---|---|---|---|
Support chatbot 100,000 replies a month, ~1,500 input tokens (system prompt + history) and ~400 output tokens each. | $350.00 | $585.00 | Claude Haiku 4.5 40% cheaper |
Document Q&A (RAG) 50,000 questions a month with ~8,000 tokens of retrieved context and ~600 output tokens each. | $550.00 | $870.00 | Claude Haiku 4.5 37% cheaper |
Coding agent 10,000 agent steps a month, ~40,000 input tokens each (70% read from the prompt cache) and ~2,000 output tokens. | $248.00 | $402.00 | Claude Haiku 4.5 38% cheaper |
Estimates use standard list prices and ignore cache-write surcharges. Different tokenizers can count the same text differently, so test with your own prompts.
Claude Haiku 4.5 vs Gemini 3.5 Flash cost calculator
- $550.00/mo$0.01 / request
- Gemini 3.5 FlashGoogle$870.00/mo$0.02 / request
Standard-tier list prices. Long-context rates apply automatically where the provider publishes a threshold. Excludes cache-write surcharges, taxes, batch and volume discounts.
Claude Haiku 4.5 pricing notes
Anthropic’s fastest model, with near-frontier intelligence.
- Cache writes cost $1.25 (5-minute cache) or $2 (1-hour cache) per 1M tokens.
- Uses Anthropic’s previous tokenizer.
Gemini 3.5 Flash pricing notes
Google’s Gemini 3.5 Flash model, with thinking, tool use and a 1M-token context window.
- Output price includes thinking tokens.
- Context cache storage costs $1.00 per 1M tokens per hour.
- Priority inference costs 1.8x the standard rate.
Frequently asked questions
- Is Claude Haiku 4.5 cheaper than Gemini 3.5 Flash?
- Claude Haiku 4.5 is 33% cheaper on input tokens ($1.00 vs $1.50 per 1M). Claude Haiku 4.5 is 44% cheaper on output tokens ($5.00 vs $9.00 per 1M). On our document Q&A workload (50,000 questions a month), Claude Haiku 4.5 costs $550.00 a month versus $870.00 for Gemini 3.5 Flash, 37% less.
- Which is more capable, Claude Haiku 4.5 or Gemini 3.5 Flash?
- Gemini 3.5 Flash scores 154.7 on the Epoch Capabilities Index vs 142.4 for Claude Haiku 4.5.
- How much does Claude Haiku 4.5 cost per 1M tokens?
- Claude Haiku 4.5 costs $1.00 per 1M input tokens and $5.00 per 1M output tokens, and $0.10 per 1M cached input tokens on Anthropic’s standard tier.
- How much does Gemini 3.5 Flash cost per 1M tokens?
- Gemini 3.5 Flash costs $1.50 per 1M input tokens and $9.00 per 1M output tokens, and $0.15 per 1M cached input tokens on Google’s standard tier.
- Which has the larger context window, Claude Haiku 4.5 or Gemini 3.5 Flash?
- Gemini 3.5 Flash has the larger context window: 1.05M tokens versus 200K for Claude Haiku 4.5.