<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
  <channel>
    <title>DEV Community: zoe wu</title>
    <description>The latest articles on DEV Community by zoe wu (@zoe_wu_e9541be3f10ed00b5c).</description>
    <link>https://dev.to/zoe_wu_e9541be3f10ed00b5c</link>
    <image>
      <url>https://media2.dev.to/dynamic/image/width=90,height=90,fit=cover,gravity=auto,format=auto/https:%2F%2Fdev-to-uploads.s3.us-east-2.amazonaws.com%2Fuploads%2Fuser%2Fprofile_image%2F4124115%2Fd2f897e5-6bec-486c-8107-eb994c159526.png</url>
      <title>DEV Community: zoe wu</title>
      <link>https://dev.to/zoe_wu_e9541be3f10ed00b5c</link>
    </image>
    <atom:link rel="self" type="application/rss+xml" href="https://dev.to/feed/zoe_wu_e9541be3f10ed00b5c"/>
    <language>en</language>
    <item>
      <title>I tracked 400+ LLM API prices every day for a month. Here's what actually changed.</title>
      <dc:creator>zoe wu</dc:creator>
      <pubDate>Mon, 14 Sep 2026 08:09:38 +0000</pubDate>
      <link>https://dev.to/zoe_wu_e9541be3f10ed00b5c/i-tracked-400-llm-api-prices-every-day-for-a-month-heres-what-actually-changed-4041</link>
      <guid>https://dev.to/zoe_wu_e9541be3f10ed00b5c/i-tracked-400-llm-api-prices-every-day-for-a-month-heres-what-actually-changed-4041</guid>
      <description>&lt;p&gt;LLM pricing pages on the web are almost always stale. Providers reprice, deprecate, and launch models weekly, and most comparison posts are frozen the day they're published.&lt;/p&gt;

&lt;p&gt;I got tired of it, so I built a tracker: every day at 03:17 UTC, a GitHub Action pulls the full OpenRouter catalog, diffs it against the previous day, and commits any price changes to a ledger. The result is &lt;a href="https://www.llmbill.app/price-history" rel="noopener noreferrer"&gt;LLMBill's price history&lt;/a&gt; — 445 models tracked, ~380 currently listed, every repricing event since July 21, 2026.&lt;/p&gt;

&lt;p&gt;A month of data turned out to be more interesting than I expected.&lt;/p&gt;

&lt;h2&gt;
  
  
  August in numbers
&lt;/h2&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;251 repricing events&lt;/strong&gt; in August alone (across ~380 live models)&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;129 price drops&lt;/strong&gt; vs &lt;strong&gt;99 hikes&lt;/strong&gt; (≥1% threshold, to filter rounding noise)&lt;/li&gt;
&lt;li&gt;Most active vendors: DeepSeek (54 events) and Qwen (54), followed by Z.ai (27) and Moonshot (18). OpenAI repriced only 13 times — frontier prices are sticky; the budget tier is where the knife fight is.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2&gt;
  
  
  The biggest drops
&lt;/h2&gt;

&lt;div class="table-wrapper-paragraph"&gt;&lt;table&gt;
&lt;thead&gt;
&lt;tr&gt;
&lt;th&gt;Model&lt;/th&gt;
&lt;th&gt;Input $/1M, Aug 1 → now&lt;/th&gt;
&lt;th&gt;Net&lt;/th&gt;
&lt;/tr&gt;
&lt;/thead&gt;
&lt;tbody&gt;
&lt;tr&gt;
&lt;td&gt;Ling 3.0 Flash&lt;/td&gt;
&lt;td&gt;$0.075 → $0.021&lt;/td&gt;
&lt;td&gt;&lt;strong&gt;−72%&lt;/strong&gt;&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;Nemotron 3 Super&lt;/td&gt;
&lt;td&gt;$0.30 → $0.085&lt;/td&gt;
&lt;td&gt;&lt;strong&gt;−72%&lt;/strong&gt;&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;DeepSeek V4 Flash (0731)&lt;/td&gt;
&lt;td&gt;$0.14 → $0.045&lt;/td&gt;
&lt;td&gt;&lt;strong&gt;−68%&lt;/strong&gt;&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;GPT-5.6 Sol (batch)&lt;/td&gt;
&lt;td&gt;$2.50 → $1.00&lt;/td&gt;
&lt;td&gt;&lt;strong&gt;−60%&lt;/strong&gt;&lt;/td&gt;
&lt;/tr&gt;
&lt;/tbody&gt;
&lt;/table&gt;&lt;/div&gt;

&lt;p&gt;The pattern: batch variants and budget/efficiency models get repriced aggressively. If your workload tolerates batch latency, August alone would have cut some bills by more than half.&lt;/p&gt;

&lt;h2&gt;
  
  
  The hikes nobody posts about
&lt;/h2&gt;

&lt;p&gt;Price drops get launch posts. Hikes happen silently:&lt;/p&gt;

&lt;div class="table-wrapper-paragraph"&gt;&lt;table&gt;
&lt;thead&gt;
&lt;tr&gt;
&lt;th&gt;Model&lt;/th&gt;
&lt;th&gt;Input $/1M&lt;/th&gt;
&lt;th&gt;Net&lt;/th&gt;
&lt;/tr&gt;
&lt;/thead&gt;
&lt;tbody&gt;
&lt;tr&gt;
&lt;td&gt;Llama 3.3 70B Instruct&lt;/td&gt;
&lt;td&gt;$0.13 → $0.91&lt;/td&gt;
&lt;td&gt;&lt;strong&gt;+610%&lt;/strong&gt;&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;Qwen3.6 27B&lt;/td&gt;
&lt;td&gt;+108%&lt;/td&gt;
&lt;td&gt;&lt;/td&gt;
&lt;/tr&gt;
&lt;tr&gt;
&lt;td&gt;Gemini 3.7 Flash&lt;/td&gt;
&lt;td&gt;+100% (doubled)&lt;/td&gt;
&lt;td&gt;&lt;/td&gt;
&lt;/tr&gt;
&lt;/tbody&gt;
&lt;/table&gt;&lt;/div&gt;

&lt;p&gt;That +610% is mostly the free/discounted era ending for a legacy model — but if your pipeline hard-coded last year's price, that's a real incident waiting to happen.&lt;/p&gt;

&lt;h2&gt;
  
  
  What I took away as a builder
&lt;/h2&gt;

&lt;ol&gt;
&lt;li&gt;
&lt;strong&gt;Budget-tier prices are volatile; frontier prices are not.&lt;/strong&gt; GPT-5-class and Claude Opus-class prices barely moved all month. The 10–50× cheaper tier is where you win — and where you must watch for silent hikes.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;"Set and forget" model routing is a bug.&lt;/strong&gt; A quarterly review of your model choice is now a cost-control measure, not a nice-to-have.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Cached input pricing changes the math entirely&lt;/strong&gt; (up to ~90% off repeated prefixes) — and cache prices reprice too.&lt;/li&gt;
&lt;/ol&gt;

&lt;h2&gt;
  
  
  How it's built
&lt;/h2&gt;

&lt;p&gt;The whole thing is boring technology, deliberately: a 40-line Node script runs in GitHub Actions daily, diffs the OpenRouter API against the committed ledger, and pushes. Vercel rebuilds the static site on push, so the public numbers are never more than a day old. No database, no server, no cron box — git &lt;em&gt;is&lt;/em&gt; the database.&lt;/p&gt;

&lt;p&gt;The live boards: &lt;a href="https://www.llmbill.app/price-history" rel="noopener noreferrer"&gt;this week's price changes&lt;/a&gt;, plus a &lt;a href="https://www.llmbill.app/" rel="noopener noreferrer"&gt;cost calculator&lt;/a&gt; if you want to price your own usage across all ~380 models.&lt;/p&gt;

&lt;p&gt;Happy to answer questions about the pipeline. And if a provider silently doubled your favorite model's price this month — you heard it here first.&lt;/p&gt;

</description>
      <category>llm</category>
      <category>ai</category>
      <category>api</category>
      <category>webdev</category>
    </item>
  </channel>
</rss>
