explainx.ainewsletter3.5k
TrendingNewsPathwaysSkills
Pricing
explainx.ai

Upskill in AI — 16 free pathways, live workshops & bootcamps, and 50+ courses from practitioners. Plus the skills, tools, and MCP servers to practice on.

follow us

corporate training

support@explainx.ai

get started

Find your pathTake Free Evaluation

learn

pathways — start freeworkshopsbootcampscoursescertificationsmock testsexplainx universitycorporate traininglearn skills & mcp

discover

skillsmcp serversexplainx mcptoolsagentsllmsdesignsdictionaryagi trackerranks

company

aboutvisionmissionteaminstructorscommunityhackathonscareers

content

daily AI newsstate of AI — live resultsblogreleasespromptsgeneratorsresource libraryfor LLMsexplainx.ai kids

solutions

all solutionsdeveloper upskillingmarketing upskillingproduct manager upskillingleadership upskilling

newsletter · weekly

Get AI news, tools, and insights in your inbox.

supportprivacytermsdata rightshow we create contentsubmission guidelines

© 2026 AISOLO Technologies Pvt Ltd

On this page

  • TL;DR — what actually changed
  • The verified old-vs-new numbers
  • Is this really "matching GPT-5" — or press framing?
  • What's actually driving the narrower gap
  • What this means if you're building on DeepSeek right now
  • Honest limitations
  • Related on explainx.ai
← Back to blog

explainx / blog

DeepSeek V4 Prices Just Went Up — Does It Really Match GPT-5.6?

DeepSeek V4's new peak pricing is live. Verified old-vs-new rates, the real percentage increase, and whether it actually matches GPT-5.6 or Claude cost.

Aug 17, 2026·8 min read·Yash Thakker
DeepSeekAI PricingOpen Weight ModelsLLM CostsChinese AI
go deep
DeepSeek V4 Prices Just Went Up — Does It Really Match GPT-5.6?

DeepSeek's new API prices are no longer a warning — they're live. At 16:00 UTC on August 16, 2026, DeepSeek replaced its flat per-token rates with a peak/off-peak schedule that raises output pricing by as much as 371% and cache-hit input pricing by more than 1,000% on the smallest tier. Multiple outlets are calling it the largest price increase since DeepSeek's API launched, and headlines framed it as DeepSeek moving to "match GPT-5 unit costs." explainx.ai flagged this coming when DeepSeek first warned of a "significant" hike on August 6, and covered the numbers as DeepSeek published them alongside the V4 Pro 0813 launch on August 13. This piece checks the "matches GPT-5" claim against the actual verified numbers, now that the new rates are live in production.

The short answer: the gap narrowed meaningfully. It did not close.

TL;DR — what actually changed

table · 2 cols
QuestionDirect answer
When did new pricing take effect?16:00 UTC, August 16, 2026
How much did output pricing rise?Up to 371% (V4 Flash) and 355% (V4 Pro) at peak hours
How much did cache-hit input rise?Up to roughly 1,100% on the cheapest tier
Did DeepSeek say it's matching GPT-5?No — that's press framing, not DeepSeek's stated reason
DeepSeek's own stated reason"Allocate resources more reasonably" and encourage usage scheduling
Is DeepSeek still cheaper than GPT-5.6 Sol/Terra?Yes, by a wide margin
Is DeepSeek still cheaper than GPT-5.6 Luna?No — Luna is now cheaper than DeepSeek V4 Pro's peak rate
Is DeepSeek still cheaper than Claude?Yes, cheaper than both Sonnet 5 and Fable 5
Weekly digest3.5k readers

Catch up on AI

Curated AI updates on agents, skills, and MCP — delivered to your inbox. Unsubscribe anytime.

The verified old-vs-new numbers

DeepSeek's own API pricing documentation confirms the new schedule. Independent reporting from InfoWorld, Bloomberg, and Caixin Global lines up with those published figures, so this is not a single-source claim.

Official DeepSeek V4 API pricing table showing peak and off-peak rates for V4 Flash and V4 Pro effective August 16, 2026, compared against pre-increase flat rates

Source: DeepSeek's official API pricing notice. Peak hours are 01:00–04:00 and 06:00–10:00 UTC; all other hours are off-peak, billed at half the peak rate.

DeepSeek V4 Flash (per million tokens):

table · 5 cols
Token typeBefore Aug 16New off-peakNew peakPeak increase
Input, cache hit$0.0028$0.007$0.014+400%
Input, cache miss$0.14$0.22$0.44+214%
Output$0.28$0.66$1.32+371%

DeepSeek V4 Pro (per million tokens):

table · 5 cols
Token typeBefore Aug 16New off-peakNew peakPeak increase
Input, cache hit$0.003625$0.022$0.044+1,114%
Input, cache miss$0.435$0.66$1.32+203%
Output$0.87$1.98$3.96+355%

Two things are worth being precise about. First, "off-peak is 50% cheaper" only describes the relationship between the two new rates — off-peak V4 Pro output at $1.98/M is still more than double the pre-August-16 flat rate of $0.87/M, not a discount off it. Second, the cache-hit tier posts the largest percentage jump by far, because it started from an unusually low base ($0.003625/M for V4 Pro) — the headline "up to 1,100%" figure comes specifically from that tier, not from the output or standard input rates most builders budget against.

Is this really "matching GPT-5" — or press framing?

DeepSeek's own notice does not say it is pricing to match GPT-5 or any competitor. Its stated rationale, per InfoWorld's reporting, is to "allocate resources more reasonably" and get developers to "schedule their tasks based on actual usage" — language about capacity constraints, not competitive positioning. That's consistent with the demand story explainx.ai covered on August 6: DeepSeek V4 Flash processed 8 trillion tokens in a single day on August 1, a volume spike that plausibly strained serving capacity ahead of the V4 Pro 0813 launch.

"Matching GPT-5 unit costs" is the aggregator and press framing layered on top of DeepSeek's numbers, not a claim DeepSeek made itself. Checked against the actual rates, it also doesn't hold up cleanly. Here's DeepSeek V4 Pro's new peak pricing next to every current GPT-5.6 tier and Anthropic's lineup, all per million tokens:

table · 4 cols
ModelInputOutputvs. DeepSeek V4 Pro peak
DeepSeek V4 Pro (peak, new)$1.32$3.96—
GPT-5.6 Luna$0.20$1.20DeepSeek is now more expensive
Claude Sonnet 5$2.00$10.00DeepSeek still ~2.5x cheaper
GPT-5.6 Terra$2.00$12.00DeepSeek still ~3x cheaper
GPT-5.6 Sol$5.00$30.00DeepSeek still ~7.6x cheaper
Claude Fable 5$10.00$50.00DeepSeek still ~12.6x cheaper

Rates for GPT-5.6's tiers and Claude Sonnet 5 come from explainx.ai's own coverage of OpenAI's July 30 GPT-5.6 price cuts and the Fable 5 cost comparison against Grok 4.6. The pattern that actually emerges is not "DeepSeek matched GPT-5." It's that DeepSeek's new peak rate landed between OpenAI's cheapest tier and its mid tier — genuinely more expensive than Luna, still comfortably cheaper than Terra, and nowhere close to flagship Sol or Fable 5 pricing. Flattening three OpenAI price points and two Anthropic ones into a single "GPT-5" number is exactly the kind of imprecision that makes a punchy headline and a misleading comparison at the same time.

What's actually driving the narrower gap

The honest 2026 pattern isn't unique to DeepSeek. Open-weight providers out of China — DeepSeek, plus Qwen, GLM, and Kimi K3 — built market share in 2025 and early 2026 on rates that were aggressive relative to their own compute costs, not just relative to closed-model rivals. That's a subsidized-launch pattern, not a stable equilibrium. As usage scales into the trillions of tokens per day and GPU capacity gets genuinely scarce, the economics catch up — a provider either builds out serving capacity fast enough to keep undercutting, or raises prices to manage load. DeepSeek's own stated rationale (resource allocation, usage scheduling) points squarely at the second path.

Meanwhile the closed-model side moved the other direction this year: OpenAI cut GPT-5.6 Luna 80% and Terra 20% in July, and Anthropic made Claude Sonnet 5's introductory $2/$10 pricing permanent instead of letting it step up as planned. Both moves push the cheap end of the closed-model market down at the same time DeepSeek's price is moving up — which is exactly why Luna, not Sol, is the tier DeepSeek now needs to worry about undercutting. The 2026 pricing story isn't "open beats closed" or "closed catches open." It's two curves converging from opposite directions, and the mid-tier is where they're actually meeting.

What this means if you're building on DeepSeek right now

  1. Re-run your cost-per-task math — don't reuse last month's estimate. A DeepSeek V4 Pro peak-hour bill just went up 3–4.5x on the metrics most workloads actually consume (cache-miss input and output). If your last cost projection used the pre-August-16 flat rate, it's stale.
  2. Move batchable work off peak hours. Off-peak (all hours outside 01:00–04:00 and 06:00–10:00 UTC) is half the peak rate. Evaluations, index refreshes, and non-interactive agent runs are the easiest candidates to reschedule.
  3. Stop treating "DeepSeek is cheapest" as a given — check the tier. Against GPT-5.6 Sol, Terra, Sonnet 5, and Fable 5, DeepSeek V4 is still clearly cheaper. Against GPT-5.6 Luna specifically, it no longer is. If your workload could tolerate Luna's capability level, it may now also be the cheaper option.
  4. Keep the price-performance comparison, not just price. A model that costs less per token but burns more tokens per completed task can still lose on total spend — the same lesson explainx.ai's Sonnet 5 vs GPT-5.6 Luna Max comparison found when comparing sticker price against real session cost.
  5. Check live pricing before you budget, not this post. DeepSeek has changed its pricing twice in about a month. Confirm current rates at DeepSeek's official API pricing page before committing production spend to either post's numbers.

Honest limitations

  • These are DeepSeek's list prices; actual delivered cost per task depends on your prompt structure, cache-hit ratio, and reasoning-effort tier, which this post does not measure directly.
  • Competitor pricing (GPT-5.6, Claude) reflects explainx.ai's own prior verified coverage as of publication and can change independently of DeepSeek's schedule.
  • "Biggest hike since launch" is a characterization repeated across multiple outlets by percentage size, not a number DeepSeek itself has published as a superlative.
  • Kimi K3, Qwen, and GLM pricing is referenced directionally in this piece; exact current per-token rates for those models aren't independently re-verified here — check each provider's own pricing page before comparing.

Related on explainx.ai

  • DeepSeek warns of a "significant" API price increase — no numbers yet
  • DeepSeek V4 Pro 0813 launch: Codex, Responses API, and new pricing
  • DeepSeek Flash hit 8 trillion tokens in a day — OpenCode's measurement
  • Anthropic makes Claude Sonnet 5 pricing permanent at $2/$10
  • OpenAI cuts GPT-5.6 Luna 80%, Terra 20%
  • Claude Sonnet 5 vs GPT-5.6 Luna Max: the cheaper workhorse
  • Perplexity adds Grok 4.6: does 60% cheaper really match Fable 5?
  • Databricks on managing AI coding costs at scale
  • DeepSeek V4 official release and peak/off-peak pricing

Primary sources: DeepSeek API pricing documentation · InfoWorld, "DeepSeek raises some V4 prices by more than 10x as AI demand strains capacity" · Bloomberg, "DeepSeek Increases Prices for AI Services by Multiple Times" · Caixin Global, "DeepSeek Launches V4-Pro and Raises API Prices by as Much as 1,100%"


Pricing figures reflect DeepSeek's published schedule effective 16:00 UTC, August 16, 2026, and explainx.ai's prior verified coverage of GPT-5.6 and Claude pricing as of publication on August 17, 2026. Rates change; confirm current numbers at each provider's official pricing page before budgeting production workloads. Follow @explainx_ai for updates.

Spotted something out of date? Let us know.
Yash Thakker

Written by

Yash Thakker

Yash is an AI expert with over 300K learners. Join his workshops →

Related posts

Aug 6, 2026

DeepSeek Warns of a "Significant" API Price Increase — No Numbers Yet

DeepSeek posted a notice warning developers of a "significant" upcoming API price increase, with no exact rates or dates disclosed. It follows days of reported record token volume that likely strained serving capacity — here's what it means for anyone budgeting around DeepSeek's rock-bottom rates.

May 23, 2026

DeepSeek V4-Pro locks in 75% permanent API discount: $0.435/M tokens, 20x cheaper than GPT-5.5

On May 22, 2026, DeepSeek made its 75% API discount permanent for V4-Pro. Coding and reasoning tasks now cost $60 for 200M tokens instead of $240+. Here's what changed, who wins, and whether cheap frontier models shift the competitive map.

Aug 14, 2026

Perplexity Adds Grok 4.6: Does 60% Cheaper Really Match Fable 5?

Perplexity's own developer changelog confirms Grok 4.6 landed on its Agent API in August 2026, days after SpaceXAI's launch. A widely repeated claim says it matches Claude Fable 5 at 60% lower cost — explainx.ai checked the actual per-token pricing and found the real gap against Fable 5 specifically is closer to 80-88%, with the 60% figure describing a different comparison.