<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
  <channel>
    <title>DEV Community: daniel</title>
    <description>The latest articles on DEV Community by daniel (@danielchinasz).</description>
    <link>https://dev.to/danielchinasz</link>
    <image>
      <url>https://media2.dev.to/dynamic/image/width=90,height=90,fit=cover,gravity=auto,format=auto/https:%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F4146151%2F243eee8b-4fa0-4656-be1a-4b15cfec51d0.png</url>
      <title>DEV Community: daniel</title>
      <link>https://dev.to/danielchinasz</link>
    </image>
    <atom:link rel="self" type="application/rss+xml" href="https://dev.to/feed/danielchinasz"/>
    <language>en</language>
    <item>
      <title>Your DeepSeek V4 Bill Can Swing 2x by the Hour — the Actual Math Across Four APIs</title>
      <dc:creator>daniel</dc:creator>
      <pubDate>Tue, 29 Sep 2026 03:45:05 +0000</pubDate>
      <link>https://dev.to/danielchinasz/your-deepseek-v4-bill-can-swing-2x-by-the-hour-the-actual-math-across-four-apis-e5j</link>
      <guid>https://dev.to/danielchinasz/your-deepseek-v4-bill-can-swing-2x-by-the-hour-the-actual-math-across-four-apis-e5j</guid>
      <description>&lt;blockquote&gt;
&lt;p&gt;Full disclosure up front: I run &lt;a href="https://ai-info.fastget.link" rel="noopener noreferrer"&gt;ai-info.fastget.link&lt;/a&gt;, where I track AI API pricing and publish the &lt;a href="https://ai-info.fastget.link/deepseek/v4-api-cost-calculator.html" rel="noopener noreferrer"&gt;DeepSeek V4 cost calculator&lt;/a&gt; this post is based on. All rates below were checked against official pricing pages and live dashboards on &lt;strong&gt;2 Sep 2026&lt;/strong&gt; — if you're reading this later, treat them as a snapshot, not gospel. No affiliate links anywhere.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;h2&gt;
  
  
  The pricing page is lying to you (sort of)
&lt;/h2&gt;

&lt;p&gt;DeepSeek doesn't have one price for V4. It has four, and which one you pay depends on decisions most developers never consciously make.&lt;/p&gt;

&lt;p&gt;Take &lt;strong&gt;V4 Flash&lt;/strong&gt;, the cheap workhorse model. DeepSeek's official list rate is $0.33 input / $0.99 output per million tokens — except that's not what you'll actually pay. DeepSeek runs peak/off-peak billing: off-peak hours cost ×0.67 of list ($0.22/$0.66), peak hours cost ×1.33 ($0.44/$1.32). So the same API call costs $0.22/M input at 3 AM and $0.44/M at 2 PM. Same model, same tokens, 2× apart. (The exact windows and how they map to your timezone are in &lt;a href="https://ai-info.fastget.link/deepseek/off-peak-pricing.html" rel="noopener noreferrer"&gt;my off-peak pricing breakdown&lt;/a&gt;.)&lt;/p&gt;

&lt;p&gt;And that's just DeepSeek's own portal. The same model ID is resold elsewhere:&lt;/p&gt;

&lt;div class="table-wrapper-paragraph"&gt;&lt;table&gt;
&lt;thead&gt;
&lt;tr&gt;
&lt;th&gt;Provider&lt;/th&gt;
&lt;th&gt;Flash (in/out per M)&lt;/th&gt;
&lt;th&gt;Pro (in/out per M)&lt;/th&gt;
&lt;/tr&gt;
&lt;/thead&gt;
&lt;tbody&gt;
&lt;tr&gt;
&lt;td&gt;DeepSeek official (list)&lt;/td&gt;
&lt;td&gt;$0.33 / $0.99&lt;/td&gt;
&lt;td&gt;$0.99 / $2.97&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;DeepSeek official, off-peak (×0.67)&lt;/td&gt;
&lt;td&gt;$0.22 / $0.66&lt;/td&gt;
&lt;td&gt;$0.66 / $1.98&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;DeepSeek official, peak (×1.33)&lt;/td&gt;
&lt;td&gt;$0.44 / $1.32&lt;/td&gt;
&lt;td&gt;$1.32 / $3.96&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;OpenRouter (cheapest host)&lt;/td&gt;
&lt;td&gt;$0.05 / $0.16&lt;/td&gt;
&lt;td&gt;$0.58 / $1.74&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;Novita&lt;/td&gt;
&lt;td&gt;flat rate, see dashboard&lt;/td&gt;
&lt;td&gt;flat rate, see dashboard&lt;/td&gt;
&lt;/tr&gt;
&lt;/tbody&gt;
&lt;/table&gt;&lt;/div&gt;

&lt;p&gt;Yes, you read the OpenRouter row right. The cheapest host on OpenRouter serves &lt;code&gt;deepseek-v4-flash-0731&lt;/code&gt; at &lt;strong&gt;$0.05/$0.16&lt;/strong&gt; — roughly a quarter of DeepSeek's off-peak rate. For Pro (&lt;code&gt;deepseek-v4-pro-0813&lt;/code&gt;), the cheapest route is Alibaba at $0.58/$1.74, which actually undercuts DeepSeek's own off-peak rate.&lt;/p&gt;

&lt;p&gt;So which number ends up on your invoice? That depends on three choices:&lt;/p&gt;

&lt;h2&gt;
  
  
  1. Official vs. reseller is a cache trade, not just a price trade
&lt;/h2&gt;

&lt;p&gt;The per-token table says "reseller wins." But look at cache pricing before you switch: DeepSeek official sells cache hits at &lt;strong&gt;$0.007–$0.044/M&lt;/strong&gt; — dramatically cheaper than anything resellers offer. If your app reuses long prompts (RAG with a fat system prompt, agents with big tool schemas, long-running chat threads), your real bill at DeepSeek official can come in far &lt;em&gt;below&lt;/em&gt; the naive per-token estimate, sometimes below the reseller's headline rate.&lt;/p&gt;

&lt;p&gt;Rule of thumb I've settled on: high cache-hit workload → official; stateless, prompt-diverse workload → the cheapest reseller route.&lt;/p&gt;

&lt;h2&gt;
  
  
  2. Peak windows are a scheduling problem, and scheduling is free
&lt;/h2&gt;

&lt;p&gt;DeepSeek's off-peak discount is the cheapest optimization in the entire AI API market, because it costs zero code quality. Batch jobs, eval runs, embeddings backfills — none of them care whether they run at 3 AM. If half your Flash tokens are batch-shaped, moving them off-peak cuts that half's cost by half ($0.22 vs $0.44 per million input — a 2× spread).&lt;/p&gt;

&lt;p&gt;The Pro model makes this more dramatic: official Pro ranges from $1.98/M output off-peak to $3.96/M at peak. Your nightly reasoning batch and your midday interactive traffic are not the same product and shouldn't share a price.&lt;/p&gt;

&lt;h2&gt;
  
  
  3. Watch the model ID, not the marketing name
&lt;/h2&gt;

&lt;p&gt;One trap worth its own section: on OpenRouter, the undated &lt;code&gt;deepseek-v4-pro&lt;/code&gt; ID routes &lt;em&gt;from $0.87/$1.74&lt;/em&gt; — above DeepSeek's own off-peak rate — while the dated &lt;code&gt;deepseek-v4-pro-0813&lt;/code&gt; starts at $0.58/$1.74 via Alibaba. Same model, different ID, different routing pool, different floor price. Pick by model ID and check which host actually serves the route; OpenRouter's "cheapest host" changes and per-route pricing isn't uniform.&lt;/p&gt;

&lt;h2&gt;
  
  
  Running your own numbers
&lt;/h2&gt;

&lt;p&gt;The &lt;a href="https://ai-info.fastget.link/deepseek/v4-api-cost-calculator.html" rel="noopener noreferrer"&gt;calculator on my site&lt;/a&gt; takes your monthly input/output token volumes and returns totals for both models across providers, with the peak/off-peak midpoint math built in. To give you scale: at a fairly typical solo-project load of 10M input / 5M output tokens per month on Flash, the spread between OpenRouter's cheapest host and DeepSeek's peak-hour official rate is the difference between roughly &lt;strong&gt;$1&lt;/strong&gt; and &lt;strong&gt;$11&lt;/strong&gt; a month. On Pro, the same spread is tens of dollars.&lt;/p&gt;

&lt;p&gt;One caveat on the table: &lt;a href="https://ai-info.fastget.link/deepseek/v4-api-cost-calculator.html" rel="noopener noreferrer"&gt;Krater.ai&lt;/a&gt; is subscription-based ($20/mo for 1,500 credits across 350+ models), so it's excluded from per-token comparisons. If your usage is spiky rather than steady, a subscription can beat all the per-token rates above — but that's a different math problem.&lt;/p&gt;

&lt;h2&gt;
  
  
  The short version
&lt;/h2&gt;

&lt;ul&gt;
&lt;li&gt;The "DeepSeek V4 price" is a range, not a number: ×0.67 off-peak to ×1.33 peak on official, and resellers undercut even the off-peak rate by 4× on Flash.&lt;/li&gt;
&lt;li&gt;Cache-heavy workloads belong on DeepSeek official ($0.007–$0.044/M cache hits); stateless workloads should shop routes.&lt;/li&gt;
&lt;li&gt;Schedule batch work into off-peak windows — it's the only API optimization that costs nothing.&lt;/li&gt;
&lt;li&gt;On OpenRouter, quote the &lt;em&gt;dated&lt;/em&gt; model ID and check the host.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Prices shift frequently on every one of these providers — I re-verify against official pages and dashboards regularly, and the calculator gets updated when they move. If a number here looks stale, the calculator page is the live copy.&lt;/p&gt;

&lt;p&gt;&lt;em&gt;Cross-posted from &lt;a href="https://ai-info.fastget.link/deepseek/v4-api-cost-calculator.html" rel="noopener noreferrer"&gt;ai-info.fastget.link&lt;/a&gt;, my site on AI API pricing and decisions — the canonical version of this post lives there.&lt;/em&gt;&lt;/p&gt;

</description>
      <category>deepseek</category>
      <category>ai</category>
      <category>api</category>
      <category>llm</category>
    </item>
  </channel>
</rss>
