July 2026 ยท DeepSeek release

DeepSeek V4 API Pricing: Flash & Pro Models

DeepSeek released V4 Flash and V4 Pro in July 2026 โ€” a significant upgrade from V3.2 with improved reasoning and 1M context. Full pricing, benchmarks, and API access details.

What is DeepSeek V4?

DeepSeek V4 is the latest model family from DeepSeek, released in July 2026 as an upgrade to V3.2. It comes in two variants โ€” Flash for cost-efficient workloads and Pro for higher capability tasks. Both models feature a 1,000,000-token context window, making them suitable for large-codebase analysis and long-document processing.

V4 retains DeepSeek's strong code generation and mathematical reasoning capabilities while improving instruction following and tool use. The model supports both text and code inputs, with a knowledge cutoff of early 2026.

Official API pricing

DeepSeek V4 Flash is one of the cheapest frontier-class models available, at $0.14 per million input tokens and $0.43 per million output tokens. V4 Pro costs about 3x more but offers higher capability. Both models use a 1M token context window.

ModelInputOutputContext
DeepSeek V4 Flash (deepseek-v4-flash)$0.14$0.431,000,000
DeepSeek V4 Pro (deepseek-v4-pro)$0.43$0.861,000,000

Source: DeepSeek official pricing as of July 2026. Pricing is per million tokens.

How V4 compares to V3.2

DeepSeek V4 Flash is priced similarly to V3.2 but offers improved reasoning across benchmarks. V4 Pro adds a higher-capability tier for tasks that need more than what Flash provides โ€” comparable to GPT-5.4 or Claude Sonnet 4.6 at a fraction of the cost.

For developers already using V3.2, upgrading to V4 Flash requires only changing the model ID from deepseek-v3.2 to deepseek-v4-flash. No API format changes are needed.

Cost comparison with other models

DeepSeek V4 Flash is among the cheapest models available with 1M context. At $0.14/$0.43 per MTok, it costs less than GPT-5.4 Mini ($0.30/$1.20), MiMo V2.5 ($0.10/$0.40), and Kimi K2.6 ($0.20/$0.80). V4 Pro at $0.43/$0.86 is still well below Claude Sonnet ($2/$10) and GPT-5.4 ($2/$8).

For high-volume coding agents and batch processing, V4 Flash offers an excellent cost-to-performance ratio. V4 Pro is a better fit when the task requires stronger multi-step reasoning or more nuanced instruction following.

API access

DeepSeek V4 models are available through the DeepSeek API and through OpenAI-compatible providers. The chat completions format is supported, so existing integrations can switch by updating the model ID. No new SDK is required.

APItokendeal provides OpenAI-compatible access to DeepSeek V4 Flash and V4 Pro through a single API key. Teams can use token packs with prepaid credits to manage costs without surprise bills.

Access DeepSeek V4 and other models

APItokendeal provides OpenAI-compatible API access to DeepSeek V4, GPT, Claude, Gemini, Kimi, and more. One API key, one dashboard, prepaid token packs with no surprise bills.