Polarison

DeepSeek V4.1 Flash vs Gemini 2.5 Pro: pricing and benchmarks

DeepSeek V4.1 Flash is 76% cheaper on input tokens ($0.30 vs $1.25 per 1M). DeepSeek V4.1 Flash is 88% cheaper on output tokens ($1.20 vs $10.00 per 1M). On our document Q&A workload (50,000 questions a month), DeepSeek V4.1 Flash costs $156.00 a month versus $800.00 for Gemini 2.5 Pro, 81% less. Epoch AI has not published a capability index score for DeepSeek V4.1 Flash yet.

Prices verified September 14, 2026.

Which should you choose?

Overall capability
Too close to call

Epoch AI has not published a capability index score for DeepSeek V4.1 Flash yet.

Price
DeepSeek V4.1 Flash

DeepSeek V4.1 Flash costs $0.525 per 1M tokens (3:1 input-to-output mix) vs $3.438 for Gemini 2.5 Pro, 85% less.

Pricing and specs

DeepSeek V4.1 FlashGemini 2.5 Pro
ProviderDeepSeekGoogle
API model IDdeepseek-flashgemini-2.5-pro
Input / 1M tokens$0.30$1.25
Cached input / 1M tokens$0.006$0.125
Output / 1M tokens$1.20$10.00
Long-context rates$2.50 in / $15.00 out above 200K prompt tokens
Context window1M tokens1.05M tokens
Max output384K tokens66K tokens
Batch discount

Benchmarks

DeepSeek V4.1 FlashGemini 2.5 Pro
Epoch Capabilities Index145.3
GPQA Diamond· Science reasoning80.4%
FrontierMath (Tiers 1–3)· Advanced math24.6%
SimpleQA Verified· Factual accuracy
ARC-AGI-2· Abstract reasoning4.9%
DeepSWE· Coding
APEX-Agents· Agentic work6.6%

Source: Epoch AI, best recorded result per model (CC BY 4.0). A dash means no published score. See all benchmarks

Monthly cost for real workloads

WorkloadDeepSeek V4.1 FlashGemini 2.5 ProDifference
Support chatbot
100,000 replies a month, ~1,500 input tokens (system prompt + history) and ~400 output tokens each.
$93.00$587.50DeepSeek V4.1 Flash 84% cheaper
Document Q&A (RAG)
50,000 questions a month with ~8,000 tokens of retrieved context and ~600 output tokens each.
$156.00$800.00DeepSeek V4.1 Flash 81% cheaper
Coding agent
10,000 agent steps a month, ~40,000 input tokens each (70% read from the prompt cache) and ~2,000 output tokens.
$61.68$385.00DeepSeek V4.1 Flash 84% cheaper

Estimates use standard list prices and ignore cache-write surcharges. Different tokenizers can count the same text differently, so test with your own prompts.

DeepSeek V4.1 Flash vs Gemini 2.5 Pro cost calculator

Standard-tier list prices. Long-context rates apply automatically where the provider publishes a threshold. Excludes cache-write surcharges, taxes, batch and volume discounts.

DeepSeek V4.1 Flash pricing notes

DeepSeek’s low-cost model with thinking and non-thinking modes, tool calls and vision input.

  • Prices shown are peak rates. Off-peak rates are half: peak hours are 01:00–04:00 and 06:00–10:00 UTC, Monday to Friday.
  • Compatible with both OpenAI-format and Anthropic-format APIs.
DeepSeek official pricing ↗

Gemini 2.5 Pro pricing notes

Google’s previous-generation Pro model.

  • Prompts over 200K tokens are billed at $2.50 input / $15 output per 1M tokens.
  • Output price includes thinking tokens.
Google official pricing ↗

Frequently asked questions

Is DeepSeek V4.1 Flash cheaper than Gemini 2.5 Pro?
DeepSeek V4.1 Flash is 76% cheaper on input tokens ($0.30 vs $1.25 per 1M). DeepSeek V4.1 Flash is 88% cheaper on output tokens ($1.20 vs $10.00 per 1M). On our document Q&A workload (50,000 questions a month), DeepSeek V4.1 Flash costs $156.00 a month versus $800.00 for Gemini 2.5 Pro, 81% less.
Which is more capable, DeepSeek V4.1 Flash or Gemini 2.5 Pro?
Epoch AI has not published a capability index score for DeepSeek V4.1 Flash yet.
How much does DeepSeek V4.1 Flash cost per 1M tokens?
DeepSeek V4.1 Flash costs $0.30 per 1M input tokens and $1.20 per 1M output tokens, and $0.006 per 1M cached input tokens on DeepSeek’s standard tier.
How much does Gemini 2.5 Pro cost per 1M tokens?
Gemini 2.5 Pro costs $1.25 per 1M input tokens and $10.00 per 1M output tokens, and $0.125 per 1M cached input tokens on Google’s standard tier.
Which has the larger context window, DeepSeek V4.1 Flash or Gemini 2.5 Pro?
Gemini 2.5 Pro has the larger context window: 1.05M tokens versus 1M for DeepSeek V4.1 Flash.

Related comparisons