Gemini 2.5 Flash vs Grok 4.6: pricing and benchmarks
Gemini 2.5 Flash is 85% cheaper on input tokens ($0.30 vs $2.00 per 1M). Gemini 2.5 Flash is 58% cheaper on output tokens ($2.50 vs $6.00 per 1M). On our document Q&A workload (50,000 questions a month), Gemini 2.5 Flash costs $195.00 a month versus $980.00 for Grok 4.6, 80% less. Grok 4.6 scores 156.3 on the Epoch Capabilities Index vs 140.5 for Gemini 2.5 Flash.
Prices verified September 14, 2026.
Which should you choose?
Grok 4.6 scores 156.3 on the Epoch Capabilities Index vs 140.5 for Gemini 2.5 Flash.
Gemini 2.5 Flash costs $0.85 per 1M tokens (3:1 input-to-output mix) vs $3.00 for Grok 4.6, 72% less.
A trade-off: Grok 4.6 is 15.8 ECI points more capable but costs 3.5x as much as Gemini 2.5 Flash.
Pricing and specs
| Gemini 2.5 Flash | Grok 4.6 | |
|---|---|---|
| Provider | xAI | |
| API model ID | gemini-2.5-flash | grok-4.6 |
| Input / 1M tokens | $0.30 | $2.00 |
| Cached input / 1M tokens | $0.03 | $0.50 |
| Output / 1M tokens | $2.50 | $6.00 |
| Long-context rates | — | $4.00 in / $12.00 out above 200K prompt tokens |
| Context window | 1.05M tokens | 500K tokens |
| Max output | 66K tokens | — |
| Batch discount | — | — |
Benchmarks
| Gemini 2.5 Flash | Grok 4.6 | |
|---|---|---|
| Epoch Capabilities Index | 140.5 | 156.3 |
| GPQA Diamond· Science reasoning | — | 92.0% |
| FrontierMath (Tiers 1–3)· Advanced math | — | 66.0% |
| SimpleQA Verified· Factual accuracy | — | 49.3% |
| ARC-AGI-2· Abstract reasoning | — | 67.1% |
| DeepSWE· Coding | — | 67.5% |
| APEX-Agents· Agentic work | 1.8% | 41.2% |
Source: Epoch AI, best recorded result per model (CC BY 4.0). A dash means no published score. See all benchmarks
Monthly cost for real workloads
| Workload | Gemini 2.5 Flash | Grok 4.6 | Difference |
|---|---|---|---|
Support chatbot 100,000 replies a month, ~1,500 input tokens (system prompt + history) and ~400 output tokens each. | $145.00 | $540.00 | Gemini 2.5 Flash 73% cheaper |
Document Q&A (RAG) 50,000 questions a month with ~8,000 tokens of retrieved context and ~600 output tokens each. | $195.00 | $980.00 | Gemini 2.5 Flash 80% cheaper |
Coding agent 10,000 agent steps a month, ~40,000 input tokens each (70% read from the prompt cache) and ~2,000 output tokens. | $94.40 | $500.00 | Gemini 2.5 Flash 81% cheaper |
Estimates use standard list prices and ignore cache-write surcharges. Different tokenizers can count the same text differently, so test with your own prompts.
Gemini 2.5 Flash vs Grok 4.6 cost calculator
- $195.00/mo$0.0039 / request
- Grok 4.6xAI$980.00/mo$0.02 / request
Standard-tier list prices. Long-context rates apply automatically where the provider publishes a threshold. Excludes cache-write surcharges, taxes, batch and volume discounts.
Gemini 2.5 Flash pricing notes
Google’s previous-generation Flash model.
- Audio input costs $1.00 per 1M tokens; text, image and video cost $0.30.
- Output price includes thinking tokens.
Grok 4.6 pricing notes
xAI’s flagship model for code and everything else, with agentic tool calling and configurable reasoning.
- Requests with 200K or more prompt tokens are billed at $4 input / $1 cached / $12 output for all tokens.
- Web Search and X Search tools cost $5 per 1,000 calls.
Frequently asked questions
- Is Gemini 2.5 Flash cheaper than Grok 4.6?
- Gemini 2.5 Flash is 85% cheaper on input tokens ($0.30 vs $2.00 per 1M). Gemini 2.5 Flash is 58% cheaper on output tokens ($2.50 vs $6.00 per 1M). On our document Q&A workload (50,000 questions a month), Gemini 2.5 Flash costs $195.00 a month versus $980.00 for Grok 4.6, 80% less.
- Which is more capable, Gemini 2.5 Flash or Grok 4.6?
- Grok 4.6 scores 156.3 on the Epoch Capabilities Index vs 140.5 for Gemini 2.5 Flash.
- How much does Gemini 2.5 Flash cost per 1M tokens?
- Gemini 2.5 Flash costs $0.30 per 1M input tokens and $2.50 per 1M output tokens, and $0.03 per 1M cached input tokens on Google’s standard tier.
- How much does Grok 4.6 cost per 1M tokens?
- Grok 4.6 costs $2.00 per 1M input tokens and $6.00 per 1M output tokens, and $0.50 per 1M cached input tokens on xAI’s standard tier.
- Which has the larger context window, Gemini 2.5 Flash or Grok 4.6?
- Gemini 2.5 Flash has the larger context window: 1.05M tokens versus 500K for Grok 4.6.