DeepSeek-V4-Flash vs Gemini 2.5 Pro: Pricing & Specs Compared

Choosing between DeepSeek-V4-Flash (DeepSeek) and Gemini 2.5 Pro (Google)? This is a side-by-side of their API pricing, context windows, and real cost on a sample workload — so you can pick on the numbers, not the marketing.

DeepSeek-V4-Flash vs Gemini 2.5 Pro: pricing & specs

Standard text/chat rates, per 1,000,000 tokens, as of June 2026. Verify against the provider's official page.
DeepSeek-V4-FlashGemini 2.5 Pro
Vendor DeepSeekGoogle
Input / 1M tokens $0.14$1.25
Output / 1M tokens $0.28$10.00
Cached input / 1M $0.0028
Context window 128K tokens1M tokens

DeepSeek-V4-Flash: Standard cache-miss input pricing; off-peak discounts may apply. Gemini 2.5 Pro: Base tier (prompts ≤200K tokens); longer prompts are billed at $2.50 input / $15 output.

Cost example: 1M input + 1M output tokens

For a workload of 1M input plus 1M output tokens, DeepSeek-V4-Flash costs $0.42 and Gemini 2.5 Pro costs $11.25 DeepSeek-V4-Flash is about 96% cheaper on this mix. Your real ratio of input to output tokens shifts the math; plug in your own numbers with the token cost calculator.

Which should you choose?

  • Lower cost: DeepSeek-V4-Flash is cheaper on output tokens ($0.28 vs $10.00 / 1M), which usually dominates spend on generation-heavy workloads.
  • Larger context window: Gemini 2.5 Pro fits more in one prompt (1M tokens) — better for long documents or big RAG context.
  • Ecosystem: pick the provider whose SDK, rate limits, and tooling fit your stack — see DeepSeek rate limits and Google rate limits.

Full specs: DeepSeek-V4-Flash · Gemini 2.5 Pro. Thinking of switching providers? See the migration guides.

Prices as of June 2026 and change often — verify before relying on them.