Gemini 3.5 Flash-Lite pricing and benchmarks
Google’s lowest-cost Gemini 3.5 model for high-volume work.
Input
$0.30
per 1M tokens
Cached input
$0.03
per 1M tokens
Output
$2.50
per 1M tokens
Capability index
145.1
#14 of 16 tracked
How capable is Gemini 3.5 Flash-Lite?
Gemini 3.5 Flash-Lite scores 145.1 on the Epoch Capabilities Index (confidence interval 142.8–146.8), ranking #14 of the 16 rated models we track.
| Benchmark | Tests | Score |
|---|---|---|
| GPQA Diamond | Science reasoning | 77.8% |
| FrontierMath (Tiers 1–3) | Advanced math | 26.0% |
| SimpleQA Verified | Factual accuracy | — |
| ARC-AGI-2 | Abstract reasoning | 10.3% |
| DeepSWE | Coding | — |
| APEX-Agents | Agentic work | — |
Source: Epoch AI (as “Gemini 3.5 Flash-Lite”), best recorded result, CC BY 4.0. A dash means no published score. Compare all models
What Gemini 3.5 Flash-Lite costs in practice
| Workload | Monthly cost |
|---|---|
Support chatbot 100,000 replies a month, ~1,500 input tokens (system prompt + history) and ~400 output tokens each. | $145.00 |
Document Q&A (RAG) 50,000 questions a month with ~8,000 tokens of retrieved context and ~600 output tokens each. | $195.00 |
Coding agent 10,000 agent steps a month, ~40,000 input tokens each (70% read from the prompt cache) and ~2,000 output tokens. | $94.40 |
Pricing details
- The same input price applies to text, image, video and audio.
- Output price includes thinking tokens.
Estimate your Gemini 3.5 Flash-Lite bill
- Gemini 3.5 Flash-LiteGoogle$195.00/mo$0.0039 / request
Standard-tier list prices. Long-context rates apply automatically where the provider publishes a threshold. Excludes cache-write surcharges, taxes, batch and volume discounts.