Polarison

Gemini 2.5 Pro vs Grok 4.6: pricing and benchmarks

Gemini 2.5 Pro is 38% cheaper on input tokens ($1.25 vs $2.00 per 1M). Grok 4.6 is 40% cheaper on output tokens ($6.00 vs $10.00 per 1M). On our document Q&A workload (50,000 questions a month), Gemini 2.5 Pro costs $800.00 a month versus $980.00 for Grok 4.6, 18% less. Grok 4.6 scores 156.3 on the Epoch Capabilities Index vs 145.3 for Gemini 2.5 Pro.

Prices verified September 14, 2026.

Which should you choose?

Overall capability
Grok 4.6

Grok 4.6 scores 156.3 on the Epoch Capabilities Index vs 145.3 for Gemini 2.5 Pro.

Advanced math
Grok 4.6

Grok 4.6 scores 66.0% on FrontierMath (Tiers 1–3) vs 24.6% for Gemini 2.5 Pro.

Price
Grok 4.6

Grok 4.6 costs $3.00 per 1M tokens (3:1 input-to-output mix) vs $3.438 for Gemini 2.5 Pro, 13% less.

Best value
Grok 4.6

Grok 4.6 delivers equal or higher overall capability for 13% less.

Pricing and specs

Gemini 2.5 ProGrok 4.6
ProviderGooglexAI
API model IDgemini-2.5-progrok-4.6
Input / 1M tokens$1.25$2.00
Cached input / 1M tokens$0.125$0.50
Output / 1M tokens$10.00$6.00
Long-context rates$2.50 in / $15.00 out above 200K prompt tokens$4.00 in / $12.00 out above 200K prompt tokens
Context window1.05M tokens500K tokens
Max output66K tokens
Batch discount

Benchmarks

Gemini 2.5 ProGrok 4.6
Epoch Capabilities Index145.3156.3
GPQA Diamond· Science reasoning80.4%92.0%
FrontierMath (Tiers 1–3)· Advanced math24.6%66.0%
SimpleQA Verified· Factual accuracy49.3%
ARC-AGI-2· Abstract reasoning4.9%67.1%
DeepSWE· Coding67.5%
APEX-Agents· Agentic work6.6%41.2%

Source: Epoch AI, best recorded result per model (CC BY 4.0). A dash means no published score. See all benchmarks

Monthly cost for real workloads

WorkloadGemini 2.5 ProGrok 4.6Difference
Support chatbot
100,000 replies a month, ~1,500 input tokens (system prompt + history) and ~400 output tokens each.
$587.50$540.00Grok 4.6 8% cheaper
Document Q&A (RAG)
50,000 questions a month with ~8,000 tokens of retrieved context and ~600 output tokens each.
$800.00$980.00Gemini 2.5 Pro 18% cheaper
Coding agent
10,000 agent steps a month, ~40,000 input tokens each (70% read from the prompt cache) and ~2,000 output tokens.
$385.00$500.00Gemini 2.5 Pro 23% cheaper

Estimates use standard list prices and ignore cache-write surcharges. Different tokenizers can count the same text differently, so test with your own prompts.

Gemini 2.5 Pro vs Grok 4.6 cost calculator

Standard-tier list prices. Long-context rates apply automatically where the provider publishes a threshold. Excludes cache-write surcharges, taxes, batch and volume discounts.

Gemini 2.5 Pro pricing notes

Google’s previous-generation Pro model.

  • Prompts over 200K tokens are billed at $2.50 input / $15 output per 1M tokens.
  • Output price includes thinking tokens.
Google official pricing ↗

Grok 4.6 pricing notes

xAI’s flagship model for code and everything else, with agentic tool calling and configurable reasoning.

  • Requests with 200K or more prompt tokens are billed at $4 input / $1 cached / $12 output for all tokens.
  • Web Search and X Search tools cost $5 per 1,000 calls.
xAI official pricing ↗

Frequently asked questions

Is Gemini 2.5 Pro cheaper than Grok 4.6?
Gemini 2.5 Pro is 38% cheaper on input tokens ($1.25 vs $2.00 per 1M). Grok 4.6 is 40% cheaper on output tokens ($6.00 vs $10.00 per 1M). On our document Q&A workload (50,000 questions a month), Gemini 2.5 Pro costs $800.00 a month versus $980.00 for Grok 4.6, 18% less.
Which is more capable, Gemini 2.5 Pro or Grok 4.6?
Grok 4.6 scores 156.3 on the Epoch Capabilities Index vs 145.3 for Gemini 2.5 Pro.
How much does Gemini 2.5 Pro cost per 1M tokens?
Gemini 2.5 Pro costs $1.25 per 1M input tokens and $10.00 per 1M output tokens, and $0.125 per 1M cached input tokens on Google’s standard tier.
How much does Grok 4.6 cost per 1M tokens?
Grok 4.6 costs $2.00 per 1M input tokens and $6.00 per 1M output tokens, and $0.50 per 1M cached input tokens on xAI’s standard tier.
Which has the larger context window, Gemini 2.5 Pro or Grok 4.6?
Gemini 2.5 Pro has the larger context window: 1.05M tokens versus 500K for Grok 4.6.

Related comparisons