<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>model-access on tomrochette.com</title>
    <link>https://tomrochette.com/tags/model-access/</link>
    <description>Recent content in model-access on tomrochette.com</description>
    <generator>Hugo -- gohugo.io</generator>
    <language>en</language>
    <managingEditor>tom@tomrochette.com (Tom Rochette)</managingEditor>
    <webMaster>tom@tomrochette.com (Tom Rochette)</webMaster>
    <copyright>© 2026 Tom Rochette</copyright>
    <lastBuildDate>Sat, 26 Sep 2026 12:43:37 -0400</lastBuildDate><atom:link href="https://tomrochette.com/tags/model-access/index.xml" rel="self" type="application/rss+xml" />
    
    <item>
      <title>Cerebras Code</title>
      <link>https://tomrochette.com/agents/model-access/cerebras-code/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/cerebras-code/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>inference</category><category>coding</category>
      <description>&lt;p&gt;Cerebras Code is a subscription from chipmaker Cerebras that sells fast inference on one open coding model at $50/month (Pro) and $200/month (Max).&#xA;&lt;strong&gt;The pitch is speed, and the record shows the speed claim and the quota fine print are the two things to verify before paying.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;A hosted coding-inference plan on Cerebras wafer-scale hardware, consumed by pointing any OpenAI-compatible editor or agent (Cline, OpenCode, Crush, Cursor) at a Cerebras API key.&#xA;It launched August 1, 2025 with Qwen3-Coder-480B advertised at up to 2,000 tokens per second and a 131k context window.&#xA;As of 2026-09-26 the product page promotes GLM 4.7 at &amp;ldquo;1,000 tokens+ per second&amp;rdquo;, so the headline model has already been swapped once.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Launched August 1, 2025 and drew 449 points and 172 comments on Hacker News the same day.&#xA;Launch windows sold out repeatedly, and as of 2026-09-26 both Pro and Max are marked &amp;ldquo;sold out&amp;rdquo; on cerebras.ai/code, with a limited free trial still open.&#xA;The model changed from Qwen3-Coder to GLM 4.7 between launch and now, which shows the plan follows whichever open model is fastest rather than committing to one family.&#xA;I could not verify funding or subscriber counts.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Speed is genuinely differentiated: even its harshest reviewer calls Cerebras the fastest provider of its model, bar none.&lt;/li&gt;&#xA;&lt;li&gt;Flat monthly pricing with published daily allowances (24M tokens Pro, 120M Max) instead of per-token anxiety.&lt;/li&gt;&#xA;&lt;li&gt;Bring-your-own-editor stance with OpenAI-compatible endpoints, no proprietary IDE lock-in.&lt;/li&gt;&#xA;&lt;li&gt;By October 2025 the original tokens-per-minute caps had been raised in response to criticism.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;The marketing number did not survive contact: InfoWorld measured well under 500 tokens/second and often under 100, against the &amp;ldquo;up to 2,000&amp;rdquo; claim.&lt;/li&gt;&#xA;&lt;li&gt;Undocumented throttles drove the experience: 300k TPM on Pro and 400k on Max produced 429 errors mid-session, and an early buyer reported a 7.5M-token daily cap hidden behind an advertised 1,000-request limit.&lt;/li&gt;&#xA;&lt;li&gt;131k context is about half the model&amp;rsquo;s native window and demands careful context management.&lt;/li&gt;&#xA;&lt;li&gt;At launch there was no prompt caching, which made agent loops expensive at the $2/1M API rate, and the usage console had defects (a Max purchase provisioning as Pro).&lt;/li&gt;&#xA;&lt;li&gt;Cerebras declined InfoWorld&amp;rsquo;s request for comment on these issues.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Pro costs $50/month with up to 24M tokens/day, and Max costs $200/month with up to 120M tokens/day.&#xA;Both plans were marked sold out as of 2026-09-26; a free tier with limited tokens remains for connection testing.&#xA;The underlying API price at launch was $2 per 1M input and $2 per 1M output on Qwen3-Coder.&#xA;The current per-token table on cerebras.ai/pricing renders client-side and I could not extract it, so treat current API rates as unverified.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2025-08-01&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pro / Max&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Launched at $50/month (24M tokens/day) and $200/month (120M tokens/day)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Cerebras blog and HN thread&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2025-09-15&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pro / Max&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;TPM caps (300k/400k) documented as the binding limit; prompt caching promised&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;InfoWorld review&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2025-10-28&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pro / Max&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Caps reported improved; Qwen3 deprecated in favor of GLM-4.6 from November&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;InfoWorld follow-up&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pro / Max&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Prices unchanged at $50/$200 but both marked sold out; model now GLM 4.7&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;cerebras.ai/code&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chutes/&#34; &gt;- Chutes&lt;/a&gt; is the cheap multi-model pay-as-you-go option; choose Chutes for price and variety, Cerebras when seconds per response matter.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;- GLM Coding Plan&lt;/a&gt; undercuts Cerebras on quota per dollar and now serves the same GLM family; choose it for volume, Cerebras for raw tokens per second.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for developers whose bottleneck is iteration latency and who code in bursts that fit the daily token allowance.&#xA;Not for heavy all-day agent runs, where the TPM throttles and 131k context cut the advertised advantage.&#xA;Claim to disagree with: at the documented throttles, Max at $200 was worse value than four Pro accounts, because 4x300k TPM beats 400k TPM.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chutes/&#34; &gt;- Chutes&lt;/a&gt; - the low-price multi-model alternative.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;- GLM Coding Plan&lt;/a&gt; - the quota-heavy alternative now running the same model family.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/cline/&#34; &gt;- Cline&lt;/a&gt; - the editor integration Cerebras documents first.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/opencode/&#34; &gt;- OpenCode&lt;/a&gt; - a terminal harness that consumes Cerebras keys.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;- Model provider feature matrix&lt;/a&gt; - where Cerebras Code sits among access providers.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.cerebras.ai/blog/introducing-cerebras-code&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.cerebras.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.cerebras.ai/blog/introducing-cerebras-code&lt;/a&gt; - launch post: $50/$200 tiers, 24M/120M tokens/day, 2,000 tok/s claim, 131k context (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.cerebras.ai/blog/qwen3-coder-480b-is-live-on-cerebras&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.cerebras.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.cerebras.ai/blog/qwen3-coder-480b-is-live-on-cerebras&lt;/a&gt; - Qwen3-Coder launch, $2/1M API rate, &amp;ldquo;20x higher coding speed&amp;rdquo; claim (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.cerebras.ai/code&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.cerebras.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.cerebras.ai/code&lt;/a&gt; - current product page: GLM 4.7 at 1,000+ tok/s, Pro and Max both marked &amp;ldquo;sold out&amp;rdquo; (fetched, HTTP 200, as of 2026-09-26).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://news.ycombinator.com/item?id=44762959&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=news.ycombinator.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://news.ycombinator.com/item?id=44762959&lt;/a&gt; - launch-day reception: 449 points, 172 comments, top comment flags missing caching (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.infoworld.com/article/4055909/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.infoworld.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.infoworld.com/article/4055909/&lt;/a&gt; - critical review, Sep 15, 2025: disputed tok/s claims, TPM caps, 131k context, billing mixups, no vendor comment (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.infoworld.com/article/4075825/how-to-vibe-code-for-free-or-almost-free.html&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.infoworld.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.infoworld.com/article/4075825/how-to-vibe-code-for-free-or-almost-free.html&lt;/a&gt; - follow-up: caps improved, Qwen3 deprecated for GLM-4.6, &amp;ldquo;fastest bar none&amp;rdquo; (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.reddit.com/r/LocalLLaMA/comments/1mfeazc/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.reddit.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.reddit.com/r/LocalLLaMA/comments/1mfeazc/&lt;/a&gt; - critical post, Aug 2, 2025: advertised 1,000 requests/day actually a 7.5M-token cap, &amp;ldquo;request&amp;rdquo; defined at ~8k tokens (retrieved, HTTP 200, via the Arctic Shift archive API).&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>ChatGPT plans</title>
      <link>https://tomrochette.com/agents/model-access/chatgpt-plans/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/chatgpt-plans/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>subscriptions</category><category>codex</category><category>openai</category>
      <description>&lt;p&gt;ChatGPT plans are OpenAI&amp;rsquo;s consumer and team subscriptions, and every tier from Free upward now carries Codex, which makes them the subscription path to OpenAI&amp;rsquo;s coding models.&lt;/p&gt;&#xA;&lt;p&gt;&lt;strong&gt;OpenAI publishes message ranges and credit rates where Anthropic publishes multipliers, but it spends the savings on complexity: six tiers, credits, speed multipliers, and a shared pool that Work, images, and voice all drain.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;The 2026 lineup: Free ($0), Go ($8/month), Plus ($20/month), Pro from $100/month in 5x and 20x versions, Business ($25/user/month, $20 annual, minimum 2 seats), and Enterprise custom.&#xA;Codex ships on web, CLI, IDE extension, and iOS, with cloud integrations (GitHub code review, Slack, Linear) from Plus up, and ChatGPT Work usage draws the same pool and rates as Codex.&#xA;The GPT-6 lineup (Astra, Sol, Luna) carries the plans; GPT-5.5 retires from ChatGPT, Work, and Codex on all plans on October 14, 2026.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Active, with the developer pricing page fetched and current as of 2026-09-26, and the widest subscription reach of any vendor here.&#xA;The 2026 lineup changed materially: Business replaced the Team plan on April 2, the same day Codex billing moved from per-message to token-based credits, and the Go tier appeared below Plus.&#xA;The current GPT-6 lineup (Astra, Sol, Luna) carries the plans, with Sol and Luna launched 2026-09-22 per the &lt;a href=&#34;https://tomrochette.com/agents/model-selection-for-coding-tasks/&#34; &gt;Model Selection guide&lt;/a&gt; and already listed as included on Plus and Pro in the fetched plan documentation.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;The most published mechanics in the category: official per-model message ranges per 5-hour window, for example GPT-6 Luna at 350-3,000 local messages on Plus versus 7,000-56,000 on Pro 20x.&lt;/li&gt;&#xA;&lt;li&gt;Credits extend usage past included limits without upgrading, and Business sits at $25/user with SSO and no training on your data, the sensible team floor.&lt;/li&gt;&#xA;&lt;li&gt;The Go tier at $8 is the cheapest named subscription carrying an agentic coder in this category.&lt;/li&gt;&#xA;&lt;li&gt;Cloud tasks open pull requests end to end, the strongest hosted-agent story among the vendor plans.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Everything shares one allowance: Codex, ChatGPT Work, image generation (3-5x faster burn), and desktop voice all draw the same pool, and cloud chats run GPT-5.6 Sol, which may consume more than local messages.&lt;/li&gt;&#xA;&lt;li&gt;A practitioner&amp;rsquo;s note from real engagements: two engineers on one shared window exhausted it by early afternoon, the April switch to credits changed budgeting for everyone.&lt;/li&gt;&#xA;&lt;li&gt;Real-world cost lands near $100-$200 per active developer per month once usage is heavy, which is Pro territory, not Plus.&lt;/li&gt;&#xA;&lt;li&gt;Speed configurations and fast mode multiply credit burn, so the same prompt costs differently depending on settings you may not remember touching.&lt;/li&gt;&#xA;&lt;li&gt;The consumer pricing page blocks automated fetches (403 this run), so plan verification runs through the developer docs.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Free: $0, Codex for quick tasks with limited usage.&#xA;Go: $8/month, lightweight coding.&#xA;Plus: $20/month, a few focused sessions a week, GPT-6 Sol and Luna included.&#xA;Pro: $100/month for 5x Plus usage, $200/month for 20x.&#xA;Business: $25/user/month ($20 annual, minimum 2 seats), Codex via workspace credits.&#xA;Enterprise: custom, flexible credit pricing.&#xA;Credits purchase past included limits; API-key Codex usage bills at API rates outside any plan.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Free to Pro&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: Free $0, Go $8, Plus $20, Pro $100 (5x) and $200 (20x) per month&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://developers.openai.com/codex/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=developers.openai.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://developers.openai.com/codex/pricing&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pro&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Split into the $100 (5x) and $200 (20x) usage tiers; a single-tier Pro preceded it&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://automationatlas.io/answers/chatgpt-codex-pricing-explained-2026&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=automationatlas.io&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://automationatlas.io/answers/chatgpt-codex-pricing-explained-2026&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-04-02&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Business&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Team at $30/user replaced by Business at $25/user ($20 annual, minimum 2 seats)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://automationatlas.io/answers/chatgpt-codex-pricing-explained-2026&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=automationatlas.io&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://automationatlas.io/answers/chatgpt-codex-pricing-explained-2026&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-04-02&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Billing&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Codex usage moved from per-message to token-based credit billing across Plus, Pro, Business, and Enterprise&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://automationatlas.io/answers/chatgpt-codex-pricing-explained-2026&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=automationatlas.io&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://automationatlas.io/answers/chatgpt-codex-pricing-explained-2026&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/claude-plans/&#34; &gt;Claude plans&lt;/a&gt;: the mirror subscription at identical price points; OpenAI publishes ranges, Anthropic publishes multipliers, and the models themselves decide more than the mechanics.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;GLM Coding Plan&lt;/a&gt;: the price-floor alternative, roughly one-seventh of Plus for the GLM line.&lt;/li&gt;&#xA;&lt;li&gt;API billing (&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model Provider Feature Matrix&lt;/a&gt;): per-token and uncapped, the right path for CI, and the escape hatch when shared windows keep interrupting real work.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for engineers inside the OpenAI ecosystem who want hosted agent features (cloud tasks, code review, Slack) with usage they can budget from published tables; Plus to start, Pro 5x when sessions interrupt.&#xA;Not for predictable maximum throughput per dollar, where flat open-model plans win.&#xA;My disagreeable claim: the published message ranges are more concrete than Anthropic&amp;rsquo;s multipliers but just as unusable in practice, because model choice, context size, and tool calls swing real consumption by an order of magnitude inside those ranges.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created when the owner asked why the Claude and OpenAI subscriptions were missing from this category.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/codex/&#34; &gt;Codex&lt;/a&gt; - the harness these plans meter and the CLI that can bypass them via API key.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/claude-plans/&#34; &gt;Claude plans&lt;/a&gt; - the rival subscription at the same price points with opposite disclosure habits.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/opencode-go/&#34; &gt;OpenCode Go&lt;/a&gt; - the flat open-model alternative at a third of Plus.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model Provider Feature Matrix&lt;/a&gt; - OpenAI&amp;rsquo;s API-side bundle, where the subscription row meets per-token reality.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://developers.openai.com/codex/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=developers.openai.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://developers.openai.com/codex/pricing&lt;/a&gt; - plan prices, per-model message ranges, credit rates, GPT-5.5 retirement, feature availability matrix (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://automationatlas.io/answers/chatgpt-codex-pricing-explained-2026&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=automationatlas.io&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://automationatlas.io/answers/chatgpt-codex-pricing-explained-2026&lt;/a&gt; - the April 2 Business replacement, the credit-billing switch, Business seat pricing, the real-world cost range and shared-window practitioner note (fetched 200, updated 2026-07-31)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.simplemetrics.xyz/chatgpt-codex-limits-2026&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.simplemetrics.xyz&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.simplemetrics.xyz/chatgpt-codex-limits-2026&lt;/a&gt; - independent limit analysis, the ranges-vs-fixed-caps framing, GPT-5.6 family tables (fetched via search extraction, updated 2026-09-09)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://help.openai.com/en/articles/11481834-chatgpt-rate-card-business-enterpriseedu&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=help.openai.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://help.openai.com/en/articles/11481834-chatgpt-rate-card-business-enterpriseedu&lt;/a&gt; - the business credit rate card and the August 31 GPT-5.4 retirement note (linked from the fetched developer pricing page; not separately fetched)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://chatgpt.com/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=chatgpt.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://chatgpt.com/pricing&lt;/a&gt; - the consumer plan page (403 to automated fetchers this run; plan prices verified through the developer docs instead)&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>Chutes</title>
      <link>https://tomrochette.com/agents/model-access/chutes/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/chutes/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>inference</category><category>pay-as-you-go</category><category>bittensor</category>
      <description>&lt;p&gt;Chutes is a model-access platform built on Bittensor (subnet 64) that sells per-token inference on open-weight models, optional monthly plans, and private GPU deployments.&#xA;&lt;strong&gt;Its per-token prices were the lowest I verified in this category, and price is the main reason to pick it.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Chutes serves open-weight models (Qwen, DeepSeek, Kimi, GLM, Mistral, Gemma) from hosted instances and charges by the million tokens, with no subscription required.&#xA;Featured models run on confidential TEE compute, and billing accepts USD or TAO from a Bittensor wallet.&#xA;A CLI can also deploy a private chute on self-serve GPU capacity billed per second, with a one-time deployment fee scaled to the GPU choice.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Active and shipping changes as of September 2026, with a live pricing page, status page, and regular community announcements.&#xA;2026 was a repair year for the business model: on February 27 the team retired the free 200-requests-per-day Early Access perk, capped subscriptions at 5x pay-as-you-go value, and cut frontier models from the $3 Base tier after its tables showed top users extracting up to 324x their subscription price.&#xA;A March 20 update reports tokens served down 45% over six weeks while revenue per million tokens rose 37.7% since February 1, after ending roughly 10B tokens/day of free-quota serving to about 11,000 users.&#xA;I could not verify funding or traffic figures.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Lowest per-token list prices I found this run, for example DeepSeek-V3.2 at $1.00/$1.00 per 1M in/out and Mistral-Nemo at $0.0245/$0.0978.&lt;/li&gt;&#xA;&lt;li&gt;Many model families behind one OpenAI-compatible key, so switching models costs nothing.&lt;/li&gt;&#xA;&lt;li&gt;Pay-as-you-go works with no plan attached, keeping commitment near zero.&lt;/li&gt;&#xA;&lt;li&gt;TEE enclaves on featured models, a privacy posture most cheap providers lack.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Terms changed three times in early 2026 (free tier removed, 5x cap, Base tier model cuts), so treat any quota as provisional.&lt;/li&gt;&#xA;&lt;li&gt;InfoWorld&amp;rsquo;s reviewer found performance underwhelming, called the privacy policy ambiguous, and needed a VPN workaround to sign up.&lt;/li&gt;&#xA;&lt;li&gt;The March 2026 update cites a global GPU shortage limiting inventory, so capacity depends on miner participation.&lt;/li&gt;&#xA;&lt;li&gt;The 6%/10% plan discounts apply only beyond the bundled daily quota, so heavy users are pushed back to full PAYG rates.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;PAYG is the default as of 2026-09-26: GLM-5.2 $1.25 in / $3.95 out per 1M, Kimi-K3 $3.00/$15.00, Qwen3-235B-A22B-Thinking $0.2989/$1.1957, DeepSeek-V3.2 $1.00/$1.00.&#xA;Plus costs $10/month for a bundled daily quota plus 6% off PAYG beyond it, and Pro costs $20/month for a larger quota plus 10% off.&#xA;Since February 27, 2026, every subscription is capped at 5x the equivalent PAYG value, with overflow billing at standard PAYG.&#xA;Private chutes run on an RTX Pro 6000 at $1.80/hour plus a $5.40 one-time deployment fee (3x the hourly rate).&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2025-07-31&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Subscriptions&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pricing tiers announced, &amp;ldquo;coming next Monday&amp;rdquo; (live early August 2025)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Chutes news post, July 31, 2025&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-02&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Base / Plus / Pro&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: $3 / $10 / $20 per month tiers existed&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Feb 27, 2026 community announcement usage tables&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-02-27&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Early Access&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Free 200-requests/day perk cut off from TEE models, fully retired March 15, 2026&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;same announcement&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-02-27&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;All subscriptions&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Value capped at 5x PAYG equivalent; GLM-5, Kimi K2.5, Qwen 3.5, MiniMax M2.5 removed from Base&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;same announcement&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Base&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;No longer listed on the pricing page, which shows only Plus $10, Pro $20, and Enterprise&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;chutes.ai/pricing&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/cerebras-code/&#34; &gt;- Cerebras Code&lt;/a&gt; sells raw speed on one coding model at $50/$200; choose it when latency dominates, Chutes when price dominates.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;- GLM Coding Plan&lt;/a&gt; sells a large fixed quota on one model family; choose it if you will burn the quota, Chutes for variety without commitment.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for hobbyists and cost-driven developers who want many open models cheaply and can tolerate churn in terms and capacity.&#xA;Not for teams needing SLAs, contractual stability, or closed frontier models.&#xA;Claim to disagree with: after the 5x cap, Plus and Pro make little sense for coders, because PAYG with a small prepaid balance delivers the same upside without a recurring bill.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/cerebras-code/&#34; &gt;- Cerebras Code&lt;/a&gt; - the speed-first subscription alternative.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;- GLM Coding Plan&lt;/a&gt; - the fixed-quota rival that wins on coding throughput per dollar.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/trackers-and-leaderboards/openrouter-rankings/&#34; &gt;- OpenRouter rankings&lt;/a&gt; - where provider model traffic becomes visible.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;- Model provider feature matrix&lt;/a&gt; - how access providers compare feature by feature.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://chutes.ai/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=chutes.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://chutes.ai/pricing&lt;/a&gt; - current PAYG model rates, Plus $10 / Pro $20 with 6%/10% PAYG discounts, private GPU pricing (fetched, HTTP 200, as of 2026-09-26).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://chutes.ai/news/community-announcement-february&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=chutes.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://chutes.ai/news/community-announcement-february&lt;/a&gt; - February 27, 2026 changes: Early Access retirement, 5x subscription cap, Base tier model removals, abuse tables (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://chutes.ai/news/from-volume-to-value-building-a-sustainable-ai-inference-platform-2&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=chutes.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://chutes.ai/news/from-volume-to-value-building-a-sustainable-ai-inference-platform-2&lt;/a&gt; - March 20, 2026 economics: tokens down 45%, revenue per token up 37.7%, free-tier costs (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://chutes.ai/news/coming-soon&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=chutes.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://chutes.ai/news/coming-soon&lt;/a&gt; - July 31, 2025 post announcing the first pricing tiers (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://chutes.ai/docs/cli/deploy&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=chutes.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://chutes.ai/docs/cli/deploy&lt;/a&gt; - private chute deployment mechanics and deployment fee structure (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.infoworld.com/article/4075825/how-to-vibe-code-for-free-or-almost-free.html&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.infoworld.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.infoworld.com/article/4075825/how-to-vibe-code-for-free-or-almost-free.html&lt;/a&gt; - critical third-party review: underwhelming performance, ambiguous privacy policy, signup bugs; confirms the $3 entry tier in October 2025 (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>Claude plans</title>
      <link>https://tomrochette.com/agents/model-access/claude-plans/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/claude-plans/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>subscriptions</category><category>claude-code</category><category>anthropic</category>
      <description>&lt;p&gt;Claude plans are Anthropic&amp;rsquo;s consumer subscriptions, Free, Pro, and Max, and from Pro upward they are the only way to run Claude Code without paying per token.&lt;/p&gt;&#xA;&lt;p&gt;&lt;strong&gt;The subscription is a multiplier on an unpublished quota, not a token bucket: you are buying 1x, 5x, or 20x of a number Anthropic never publishes, drawn from one shared pool that chat, desktop, and Claude Code all drain together.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Three consumer tiers from Anthropic (PBC): Free at $0 with no Claude Code, Pro at $20 per month ($17 per month billed annually at $200 up front), and Max from $100 per month in 5x and 20x versions.&#xA;Pro carries Claude Code, Claude Science, Design, Slides, Docs, and Projects; Max adds priority access at high-traffic times, higher output limits, and early access.&#xA;Team seats run $25 per month ($20 annual) for Standard and $100 ($125 monthly, $100 annual) for Premium with 5x Standard usage; Enterprise is custom.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Active and the default way engineers pay for Claude Code, with the pricing page fetched and current as of 2026-09-26.&#xA;The 2026 record is lively: an April pricing-page test briefly removed Claude Code from Pro and was reverted within a day, a caching bug behind spring &amp;ldquo;usage drain&amp;rdquo; complaints was postmortemed on April 23 with limits reset, 5-hour limits were permanently doubled on May 6, and the weekly-limit promotion of May 13 was extended through August 19, 2026.&#xA;&lt;strong&gt;The subscription has been repricing its value, not its price: the dollar figures are stable while the quota they buy keeps moving, which makes old guides the main hazard.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Flat pricing against metered fear: one widely cited power user pushed roughly 10 billion tokens through eight months, about $15,000 at Opus API rates, on a $200 Max subscription (secondary writeup, hold loosely).&lt;/li&gt;&#xA;&lt;li&gt;Usage credits continue a session past the plan limit at API rates under a cap you set, so a long task need not die at the wall.&lt;/li&gt;&#xA;&lt;li&gt;Max gets priority access at peak times, which is when Pro users feel throttling.&lt;/li&gt;&#xA;&lt;li&gt;The limits moved in your favor twice in spring 2026, permanently for the 5-hour clock.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;One shared pool: a morning of claude.ai chat shrinks the afternoon Claude Code session, and agent teams burn about 7x a normal session in plan mode, per Anthropic&amp;rsquo;s own costs docs.&lt;/li&gt;&#xA;&lt;li&gt;Fable 5.1, the flagship, is capped at 50% of weekly limits on Max and is available on Pro only through usage credits, the first tightening of 2026 (July 20).&lt;/li&gt;&#xA;&lt;li&gt;Anthropic does not publish token quotas for any plan; multipliers are all you get, and third-party figures like &amp;ldquo;220,000 tokens per 5 hours&amp;rdquo; are estimates.&lt;/li&gt;&#xA;&lt;li&gt;The weekly-limit promotion lapsed on paper on August 19, 2026, so weekly walls may now arrive about a third sooner than spring habits suggest.&lt;/li&gt;&#xA;&lt;li&gt;A June 15 plan to move programmatic usage (Agent SDK, headless runs) to separate credit pools was paused the day it was due and never shipped; watch for it to return.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Free: $0, no Claude Code.&#xA;Pro: $20/month, $17/month billed annually ($200 up front), Claude Code included.&#xA;Max 5x: $100/month, 5x Pro usage.&#xA;Max 20x: $200/month, 20x Pro usage, monthly billing only.&#xA;Team Standard $25/seat/month ($20 annual), Team Premium $100/seat/month ($100 annual), Enterprise custom; usage credits extend any plan at API rates under a spending cap.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pro / Max&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: Pro $20/month ($17 annual), Max 5x $100/month, Max 20x $200/month, Fable on Pro via usage credits and on Max at 50% of weekly limits&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://claude.com/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=claude.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://claude.com/pricing&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-05-06&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Usage&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;5-hour limits permanently doubled for Pro, Max, Team, and seat-based Enterprise; peak-hour throttling removed&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://ccforeveryone.com/guides/claude-code-limits-and-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=ccforeveryone.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://ccforeveryone.com/guides/claude-code-limits-and-pricing&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-05-13&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Usage&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Weekly limits raised 50% as a promotion, extended three times through 2026-08-19&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://ccforeveryone.com/guides/claude-code-limits-and-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=ccforeveryone.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://ccforeveryone.com/guides/claude-code-limits-and-pricing&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-07-20&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Fable&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Fable 5.1 added to Max, Team Premium, and Enterprise up to 50% of weekly limits; Pro and Team Standard only via usage credits, the year&amp;rsquo;s first tightening&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://ccforeveryone.com/guides/claude-code-limits-and-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=ccforeveryone.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://ccforeveryone.com/guides/claude-code-limits-and-pricing&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-07-25&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Models&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Opus 5 released on all paid plans at unchanged subscription prices, default on Max&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://ccforeveryone.com/guides/claude-code-limits-and-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=ccforeveryone.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://ccforeveryone.com/guides/claude-code-limits-and-pricing&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;GLM Coding Plan&lt;/a&gt;: the price-floor challenger at $18, with published credit mechanics and China-based routing; Claude plans cost more and publish less, and buy the stronger frontier model.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chatgpt-plans/&#34; &gt;ChatGPT plans&lt;/a&gt;: the direct rival subscription, which publishes per-model message ranges and credit rates where Anthropic publishes multipliers only.&lt;/li&gt;&#xA;&lt;li&gt;API billing (&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model Provider Feature Matrix&lt;/a&gt;): still the right path for CI and automation, and the comparison that makes Max look cheap only works at genuinely heavy usage.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for engineers whose daily driver is Claude Code and who can live with a shared pool and unpublished quotas; start at Pro and upgrade on the second real weekly cap.&#xA;Not for automation-heavy workloads, where the API or a flat open-model plan fits better.&#xA;My disagreeable claim: most Max subscribers are paying for a productivity feeling rather than a quota, because the median engineer never hits the Pro weekly wall that justifies 5x.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created when the owner asked why the Claude and OpenAI subscriptions were missing from this category.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/claude-code/&#34; &gt;Claude Code&lt;/a&gt; - the harness these plans meter, with its own token-overhead numbers.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chatgpt-plans/&#34; &gt;ChatGPT plans&lt;/a&gt; - the rival subscription, more published mechanics at the same price points.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;GLM Coding Plan&lt;/a&gt; - the price-floor alternative for the same daily-driver role.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model Provider Feature Matrix&lt;/a&gt; - Anthropic&amp;rsquo;s API-side bundle, where the subscription row meets per-token reality.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://claude.com/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=claude.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://claude.com/pricing&lt;/a&gt; - plan prices, annual discount, Claude Code inclusion matrix, Fable access rules (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://ccforeveryone.com/guides/claude-code-limits-and-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=ccforeveryone.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://ccforeveryone.com/guides/claude-code-limits-and-pricing&lt;/a&gt; - the 2026 change log (May 6 doubling, May 13 promotion, July 20 Fable tightening, paused Agent SDK split), limit mechanics, Team seat prices, the API-versus-subscription math (fetched 200, updated 2026-08-05)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.ssdnodes.com/blog/claude-code-pricing-in-2026-every-plan-explained-pro-max-api-teams/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.ssdnodes.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.ssdnodes.com/blog/claude-code-pricing-in-2026-every-plan-explained-pro-max-api-teams/&lt;/a&gt; - independent plan walkthrough, the instrumented-usage cost comparison it cites (fetched via search extraction, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.layer3labs.io/guides/claude-pro-vs-max-for-teams&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.layer3labs.io&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.layer3labs.io/guides/claude-pro-vs-max-for-teams&lt;/a&gt; - the weekly-cap skepticism and the upgrade heuristics (fetched via search extraction, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://support.claude.com/en/articles/11049741-what-is-the-max-plan&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=support.claude.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://support.claude.com/en/articles/11049741-what-is-the-max-plan&lt;/a&gt; - the Max plan article: $100/$200 tiers, monthly-only billing, 5-hour and weekly limit mechanics, the discretionary-limits reservation (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>GLM Coding Plan</title>
      <link>https://tomrochette.com/agents/model-access/glm-coding-plan/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/glm-coding-plan/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>subscription</category><category>coding</category>
      <description>&lt;p&gt;The GLM Coding Plan is Z.ai&amp;rsquo;s subscription that sells access to the GLM model family inside coding agents, currently $18/$80/$168 per month for Lite/Pro/Max.&#xA;&lt;strong&gt;It is the cheapest quota-per-dollar coding subscription I verified, but a multiplier system governs the quota, so the sticker price is only the start of the math.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;A credit-based subscription from Z.ai (Zhipu) for GLM-5.3 and GLM-5.3-Flash, consumed inside supported coding tools (Claude Code, Cline, OpenCode, Kilo Code, Cursor, Z.ai&amp;rsquo;s own ZCode, and others).&#xA;You point the tool at Z.ai&amp;rsquo;s Anthropic- or OpenAI-compatible endpoints with a plan key, and requests to older GLM versions are auto-routed to 5.3.&#xA;Plans bundle exclusive MCP servers (vision, web search, web reader, Zread).&#xA;Usage outside officially supported tools is contractually prohibited, which makes this a walled-garden subscription rather than an API credit pack.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Actively developed and restructured twice in a year: prompt-based plans at launch in 2025, then a credits-based system for new subscribers on July 30, 2026, with legacy plans grandfathered to the end of their billing cycle.&#xA;The plan rode the GLM release cadence (GLM-4.5 in 2025, GLM-5.2 in June 2026, GLM-5.3 in August 2026), and tier prices rose at each step.&#xA;Promotional quota mechanics (off-peak discounts, Flash campaigns) changed repeatedly through 2026, so published value figures have a short shelf life.&#xA;I found no reliable subscriber counts.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Quota per dollar is unmatched in this category: Z.ai itself documents quota worth approximately 15-30x the monthly fee at API rates.&lt;/li&gt;&#xA;&lt;li&gt;Lite at $18, or $12.60 effective on yearly billing, undercuts every $20 rival I checked.&lt;/li&gt;&#xA;&lt;li&gt;Off-peak usage (outside Mon-Fri 14:00-18:00 UTC+8) deducts credits at 50%, a standing discount for most time zones.&lt;/li&gt;&#xA;&lt;li&gt;Hard-stop quota means spend is fixed: when quota runs out, calls pause and your account balance is never touched.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Subscriptions are non-refundable once purchased, per the official FAQ.&lt;/li&gt;&#xA;&lt;li&gt;The flagship model burns quota up to 3x faster at peak (a usage multiplier, not a price change), so effective throughput can be a third of the headline prompts.&lt;/li&gt;&#xA;&lt;li&gt;Strict supported-tools-only terms and special endpoints block custom integrations.&lt;/li&gt;&#xA;&lt;li&gt;Third-party reviewers report slow and unreliable periods under load, and prompts route through China-based infrastructure, a blocker for residency-bound teams.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Current tiers as of 2026-09-26: Lite $18/month (per Z.ai docs), Pro $80 and Max $168 per August 2026 snapshots of the subscribe page, with 20% off quarterly and 30% off yearly billing (effective $12.60/$56/$117.60 per month).&#xA;The July 2026 credits system allocates 2,000/10,000 credits per 5 hours/week on Lite, 12,000/60,000 on Pro, 28,000/140,000 on Max.&#xA;Credits deduct per token type with multipliers (GLM-5.3: input 6.9, cached input 1.7, output 24, over 10,000), plus per-call charges for bundled MCP tools.&#xA;Under the legacy prompts system, Z.ai documented quota as roughly 80/400/1,600 prompts per 5-hour window, with one prompt estimated at 15-20 model invocations.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2025-09&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Lite / Pro&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: launched at $6/$30 per month with first-month promos of $3/$15&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Cline blog, Sep 18, 2025&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-06-18&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Lite / Pro / Max&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;List prices $18/$72/$160, yearly effective $12.60/$50.40/$112&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;HyScaler&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-07-30&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;All plans&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Credits-based plans replace prompt-based plans for new subscribers&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;docs.z.ai usage-revision notice&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-08-14&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Lite / Pro / Max&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pro to $80 and Max to $168; 20% quarterly and 30% yearly discounts&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;emergent.sh&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/cerebras-code/&#34; &gt;- Cerebras Code&lt;/a&gt; sells speed at $50/$200; choose it when latency dominates, GLM when volume per dollar dominates.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chutes/&#34; &gt;- Chutes&lt;/a&gt; sells cheap pay-as-you-go across many model families; choose it for flexibility, GLM for a large fixed quota on one frontier family.&lt;/li&gt;&#xA;&lt;li&gt;Z.ai&amp;rsquo;s own &lt;a href=&#34;https://tomrochette.com/agents/harnesses/zcode/&#34; &gt;ZCode&lt;/a&gt; is the intended surface, but any supported harness works with the same quota.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for high-volume, human-in-the-loop coding where quota per dollar is the binding constraint and you can work off-peak.&#xA;Not for long-horizon autonomous agent runs (aggregator tracking puts GLM-5.2 at about half of Claude Opus 4.8&amp;rsquo;s SWE-Marathon score), Claude-ecosystem power users, or residency-bound teams.&#xA;Claim to disagree with: the rational entry is a single month of $18 Lite and never the discounted yearly plan, because the plan is non-refundable and its terms changed twice in a year.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/cerebras-code/&#34; &gt;- Cerebras Code&lt;/a&gt; - the speed-first alternative subscription.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chutes/&#34; &gt;- Chutes&lt;/a&gt; - the cheap multi-model alternative.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/zcode/&#34; &gt;- ZCode&lt;/a&gt; - Z.ai&amp;rsquo;s own coding surface for this plan.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/cline/&#34; &gt;- Cline&lt;/a&gt; - one of the first tools the plan targeted.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-selection-for-coding-tasks/&#34; &gt;- Model selection for coding tasks&lt;/a&gt; - how to decide whether GLM models fit your tasks.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://docs.z.ai/devpack/overview&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=docs.z.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://docs.z.ai/devpack/overview&lt;/a&gt; - current credit allowances (2,000/12,000/28,000 per 5h), credit formula, off-peak 50% rule, &amp;ldquo;starting at 18 USD&amp;rdquo; (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://docs.z.ai/devpack/notice/usage-revision&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=docs.z.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://docs.z.ai/devpack/notice/usage-revision&lt;/a&gt; - July 30, 2026 credits transition; legacy 15-30x value claim and per-plan prompt ceilings (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://docs.z.ai/devpack/faq&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=docs.z.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://docs.z.ai/devpack/faq&lt;/a&gt; - non-refundable policy, supported-tools restriction, no balance deduction, quota reset cards (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://cline.bot/blog/zai-cline-3-dollar-ai-coding&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=cline.bot&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://cline.bot/blog/zai-cline-3-dollar-ai-coding&lt;/a&gt; - launch-era pricing: $6/$30 with $3/$15 first-month promos, 120/600 prompts per 5 hours (fetched, HTTP 200, Sep 18, 2025).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://hyscaler.com/insights/glm-coding-plan-review/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=hyscaler.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://hyscaler.com/insights/glm-coding-plan-review/&lt;/a&gt; - June 2026 tiers at $18/$72/$160, 3x peak multiplier, competitive framing (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.digitalapplied.com/blog/glm-coding-plan-worth-it-2026-value-analysis&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.digitalapplied.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.digitalapplied.com/blog/glm-coding-plan-worth-it-2026-value-analysis&lt;/a&gt; - critical analysis, Jul 3, 2026: multiplier math, non-refundability, China residency, long-horizon benchmark gap (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://emergent.sh/learn/glm-5-3-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=emergent.sh&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://emergent.sh/learn/glm-5-3-pricing&lt;/a&gt; - August 2026 tiers at $18/$80/$168 with yearly equivalents; GLM-5.3 has no per-token API rate yet (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>Google AI plans</title>
      <link>https://tomrochette.com/agents/model-access/google-ai-plans/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/google-ai-plans/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>google</category><category>subscriptions</category><category>gemini</category>
      <description>&lt;p&gt;Google AI plans are Google&amp;rsquo;s consumer subscription ladder, Google AI Plus, Google AI Pro, and Google AI Ultra in 5x and 20x versions, and from Plus upward they are how individuals raise the limits on Antigravity, AI Studio, Jules, and the Gemini app.&lt;/p&gt;&#xA;&lt;p&gt;&lt;strong&gt;The plans sell multipliers on unpublished quotas plus a credit pool billed at API consumption rates, and Google&amp;rsquo;s own free Antigravity tier is good enough that most engineers hitting a wall should check whether they need to pay at all.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Three consumer tiers sold through Google One: Plus (2x Gemini access versus non-subscribers, 400 GB storage), Pro (4x, 5 TB, $10 in monthly Google Cloud credits), and Ultra in 5x (5x Pro, 20 TB, $40 credits) and 20x (20x Pro, 30 TB, $100 credits) versions, as of 2026-09-26.&#xA;The plans carry the Gemini app, Google Flow, Gemini Notebook, AI Studio, Android Studio, Chrome auto browse, and expanded Antigravity and Jules limits, with Flow credits of 200 to 25,000 per month across the ladder.&#xA;Personal Google Accounts only: Workspace customers are pushed to Gemini add-ons, and organizations to Google Cloud terms with consumption-based API pricing.&#xA;The FAQ confirms the ladder&amp;rsquo;s churn: the old Google AI Premium plan was renamed Google AI Plus.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Active and restructured repeatedly, with the tiers, multipliers, and storage allotments verified from Google&amp;rsquo;s own pages on 2026-09-26.&#xA;Plus is available in over 160 countries, Pro and Ultra in over 150.&#xA;The 2026 record shows a rename (AI Premium to Plus) and an Ultra split into 5x and 20x columns, both visible on the current comparison table.&#xA;&lt;strong&gt;Prices are the weak point of the product surface: the marketing pages render the dollar figure client-side per region, so my US and Canada fetches returned the ladder with the price stripped, and third-party guides are the only place the US numbers appear in text.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;The free Antigravity tier it upgrades is real: unlimited tab completions and command requests, a weekly baseline quota, and access to Gemini, Claude, and open-weight models.&lt;/li&gt;&#xA;&lt;li&gt;The flexible AI credit pool converts a hard wall into metered overage at documented consumption pricing, with a Never/Always overage setting instead of a surprise bill.&lt;/li&gt;&#xA;&lt;li&gt;Ultra refreshes its quota every five hours against the free tier&amp;rsquo;s weekly clock, the difference that matters for daily-driver agent work.&lt;/li&gt;&#xA;&lt;li&gt;The bundled Google Cloud credits ($10, $40, $100 monthly by tier) partially self-fund Ultra for anyone who also calls the Gemini API.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Google publishes no quota numbers for any tier, and the Antigravity docs say limits are &amp;ldquo;correlated with the amount of work done by the agent&amp;rdquo;, so one gnarly prompt can cost a day of quota.&lt;/li&gt;&#xA;&lt;li&gt;The AI credits purchase page sits behind Google sign-in, so credit prices are the least documented numbers in the whole stack.&lt;/li&gt;&#xA;&lt;li&gt;Google&amp;rsquo;s own pages disagree on Plus storage, 400 GB on the AI plans page versus 2 TB on the Canadian plans page, both fetched 2026-09-26.&lt;/li&gt;&#xA;&lt;li&gt;The consumption rates your credits burn at double on January 1, 2027, when the current Gemini API promotional prices expire.&lt;/li&gt;&#xA;&lt;li&gt;Tier mechanics have already been changed out from under users once: the May 2026 &amp;ldquo;Antigravity bait and switch&amp;rdquo; thread drew 771 points.&lt;/li&gt;&#xA;&lt;li&gt;No bring-your-own-key and no organizational contract tiers on the consumer path.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Canadian storefront, fetched 2026-09-26: Google AI Plus CA$13.99/month, Google AI Pro CA$26.99/month, Ultra not listed on that page.&#xA;US dollar list prices did not render in my fetched HTML; a third-party comparison table lists Plus around $8/month as the cheapest paid path and Ultra 20x at $200/month as the top individual tier (blocked fetch, see references).&#xA;Organizations: Google Cloud terms with consumption-based API pricing, included in select Gemini Enterprise subscriptions.&#xA;Credits bill at standard Gemini Enterprise consumption rates, anchored by the Gemini API price list: Gemini 3.8 Flash at $0.75/$3.75 per million tokens through December 31, 2026, doubling in 2027, and Gemini 3.1 Pro at $2.00/$12.00.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-08-24&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Individuals / Pro / Ultra&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: Antigravity free at $0 with basic weekly rate limits, Google AI Pro and Ultra the paid paths that raise limits and add the AI credit pool&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://antigravity.google/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=antigravity.google&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://antigravity.google/pricing&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Ladder&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Verified: Plus / Pro / Ultra (5x, 20x) ladder at 2x/4x/5x-Pro/20x-Pro access, AI Premium confirmed renamed to Google AI Plus, CA$13.99 Plus and CA$26.99 Pro on the Canadian storefront&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://one.google.com/about/google-ai-plans&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=one.google.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://one.google.com/about/google-ai-plans&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/claude-plans/&#34; &gt;Claude plans&lt;/a&gt;: the same multiplier-on-a-secret-quota pattern, but Google bundles harder (Flow credits, cloud credits, YouTube) and prices its entry tier lower.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chatgpt-plans/&#34; &gt;ChatGPT plans&lt;/a&gt;: the rival ladder, which publishes per-model message ranges where Google publishes adjectives.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/surfaces/antigravity/&#34; &gt;Google Antigravity&lt;/a&gt;: the surface these plans meter; its free tier undercuts the paid ladder for light use, which is the sharpest critique of the ladder itself.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model Provider Feature Matrix&lt;/a&gt;: the API-side alternative, still the right answer for automation and CI, and the baseline your credit pool burns at.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for engineers who have actually hit the free Antigravity weekly wall and want metered overage at known API rates; start at Pro, and only after two real capped weeks.&#xA;Not for light users, who should stay on the free tier, and not for automation-heavy workloads, which belong on the API.&#xA;My disagreeable claim: Ultra is priced for video generation more than for coding, because the 25,000 monthly Flow credits are its most concrete differentiator, so engineers buying Ultra 20x for Antigravity are mostly buying Flow.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created when the owner asked for any remaining subscription providers.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/claude-plans/&#34; &gt;Claude plans&lt;/a&gt; - the multiplier-based subscription this ladder most resembles, with a longer 2026 change log.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chatgpt-plans/&#34; &gt;ChatGPT plans&lt;/a&gt; - the OpenAI ladder at the same price points, with more published mechanics.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/surfaces/antigravity/&#34; &gt;Google Antigravity&lt;/a&gt; - the free-first editor and agent platform whose limits these plans raise.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model Provider Feature Matrix&lt;/a&gt; - where the subscription ends and per-token consumption pricing begins.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://one.google.com/about/google-ai-plans&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=one.google.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://one.google.com/about/google-ai-plans&lt;/a&gt; - the plan ladder, comparison table (multipliers, Flow credit quotas, Antigravity and Jules rows), AI Premium rename, country counts (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://one.google.com/intl/en_us/about/google-ai-plans/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=one.google.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://one.google.com/intl/en_us/about/google-ai-plans/&lt;/a&gt; - US variant: $10/$40/$100 monthly Google Cloud credits by tier, Home Premium bundle values, Ultra 5x and 20x storage tiers (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://one.google.com/about/plans&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=one.google.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://one.google.com/about/plans&lt;/a&gt; - Canadian storefront prices, CA$13.99 Plus and CA$26.99 Pro, no Ultra listed (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://antigravity.google/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=antigravity.google&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://antigravity.google/pricing&lt;/a&gt; - the $0 individual tier, Pro/Ultra as the paid paths adding rate limits and the AI credit pool, Google Cloud consumption terms for organizations (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://antigravity.google/docs/plans&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=antigravity.google&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://antigravity.google/docs/plans&lt;/a&gt; - quota mechanics: five-hour Ultra refresh, weekly free quota, credits billed at standard Gemini Enterprise consumption pricing, Never/Always overages, no BYOK (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://ai.google.dev/gemini-api/docs/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=ai.google.dev&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://ai.google.dev/gemini-api/docs/pricing&lt;/a&gt; - the consumption-rate anchor: 3.8 Flash $0.75/$3.75 per million tokens until 2026-12-31 then doubled, 3.1 Pro $2.00/$12.00 (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://news.ycombinator.com/item?id=48222529&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=news.ycombinator.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://news.ycombinator.com/item?id=48222529&lt;/a&gt; - critical source: the May 2026 &amp;ldquo;Antigravity bait and switch&amp;rdquo; thread, 771 points, on tiers changing under existing users (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://codeagentswarm.com/en/guides/grok-build-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=codeagentswarm.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://codeagentswarm.com/en/guides/grok-build-pricing&lt;/a&gt; - third-party table naming Plus around $8/month and Ultra 20x at $200/month; fetched 429 this run, figures carried from the 2026-08-25 verification, hold loosely&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>Kimi Code</title>
      <link>https://tomrochette.com/agents/model-access/kimi-code/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/kimi-code/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>coding-subscription</category><category>moonshot-ai</category><category>pricing</category>
      <description>&lt;p&gt;Kimi Code is Moonshot AI&amp;rsquo;s developer coding subscription, selling the K-series models through a CLI, desktop app, and VS Code extension under a membership quota.&#xA;&lt;strong&gt;The same membership is sold at a yen ladder on kimi.com and a dollar ladder on kimi.ai, and the entry tiers do not convert between them.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Moonshot AI sells Kimi Code as the developer service inside a Kimi membership, with CLI, desktop, IDE, and third-party requests drawing on one shared quota.&#xA;Third-party harnesses connect through OpenAI-compatible and Anthropic-compatible endpoints on api.kimi.com.&#xA;Model access is tiered: K2.7 Code for every paying member, K3 at full 1M context from the mid tier up, with RMB-billed Extra Usage as overflow.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Active and restructuring quickly: K3 shipped into Kimi Code on July 17, 2026, and by late September a restructure had renamed tiers and dropped the weekly quota window for new members.&#xA;The CLI is published on GitHub under MoonshotAI, with third-party coverage putting it at 3.2k stars in July 2026.&#xA;Independent guides multiplied through 2026, which I read as real adoption.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;The $19 entry ($15 billed annually) undercuts Claude Pro and Cursor Pro while including a 1M-context frontier-class model at the mid tier.&lt;/li&gt;&#xA;&lt;li&gt;The yen ladder starts at ¥49 per month, well below the $19 international entry at any recent exchange rate.&lt;/li&gt;&#xA;&lt;li&gt;Anthropic-compatible endpoints mean Claude Code and OpenCode work without new tooling.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Coding quota shares one pool with chat, Deep Research, and other Kimi features, so non-coding usage drains your coding allowance.&lt;/li&gt;&#xA;&lt;li&gt;The plan ladder moved twice in 2026, so quota rules at signup may not persist.&lt;/li&gt;&#xA;&lt;li&gt;A February 2026 checkout experiment offered personalized first months between $0.99 and $11.99, so two subscribers on one tier can pay different prices.&lt;/li&gt;&#xA;&lt;li&gt;One four-month user quit, reporting that higher tiers bought less inference per dollar, not more.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;On kimi.ai: Moderato $19, Allegretto $39, Allegro $99, Vivace $199 per month, annual effective $15/$31/$79/$159 (as of 2026-09-26).&#xA;On kimi.com: Andante ¥49, Moderato ¥99, Allegretto ¥199, Allegro ¥699 per month (as of 2026-09-26).&#xA;Under the new ladder, Go has no coding quota, Plus and above include Kimi Code, and Pro and above unlock K3 at 1M context.&#xA;Extra Usage is pay-as-you-go overflow with a ¥25 minimum top-up, generally non-refundable.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-07-17&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Moderato to Vivace&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Listed $19/$39/$99/$199 with time-limited promos at $15/$31/$79/$159; Kimi and Kimi Code separation announced&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://www.digitalapplied.com/blog/kimi-code-k3-hands-on-setup-plans-cache-2026&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.digitalapplied.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.digitalapplied.com/blog/kimi-code-k3-hands-on-setup-plans-cache-2026&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Same tiers&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Promos became standard annual pricing; new plans released at unchanged prices, weekly window dropped for new members&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://www.kimi.ai/help/membership/membership-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.kimi.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.kimi.ai/help/membership/membership-pricing&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;MiniMax Coding Plan (../minimax-coding-plan/index.md) meters tokens across modalities instead of time windows, cheaper at the top end but carrying 2026 billing-change baggage.&#xA;NanoGPT (../nanogpt/index.md) fits when open-weight breadth matters more than one lab&amp;rsquo;s flagship.&#xA;Claude Code or Cursor remain the picks when reliability per hour beats price per hour.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for cost-sensitive daily coders who want K3&amp;rsquo;s 1M context at $15 to $39 per month and accept quota churn.&#xA;Not for teams needing contract-stable plan terms after two restructures in one year.&#xA;I will claim something arguable: the ¥49 China-market ladder proves this product profits at roughly a third of the international price, which makes $19 a regional tax, not a cost.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/minimax-coding-plan/&#34; &gt;MiniMax Coding Plan&lt;/a&gt; - the other first-party Chinese coding plan, token-metered where Kimi is window-metered.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/nanogpt/&#34; &gt;NanoGPT&lt;/a&gt; - the gateway alternative when model breadth beats model loyalty.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/opencode/&#34; &gt;OpenCode&lt;/a&gt; - a harness that drives Kimi Code&amp;rsquo;s Anthropic-compatible endpoint.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-selection-for-coding-tasks/&#34; &gt;Model selection for coding tasks&lt;/a&gt; - pick the model before the plan.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.kimi.com/code/en&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.kimi.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.kimi.com/code/en&lt;/a&gt; - official product page; surfaces and the K3 1M-context claim.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.kimi.com/code/docs/en/kimi-code/membership.html&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.kimi.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.kimi.com/code/docs/en/kimi-code/membership.html&lt;/a&gt; - membership docs; new ladder (Go/Plus/Pro), weekly window removal, Extra Usage terms.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.kimi.ai/help/membership/membership-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.kimi.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.kimi.ai/help/membership/membership-pricing&lt;/a&gt; - dollar pricing $19/$39/$99/$199 monthly, $15/$31/$79/$159 annual, as of 2026-09-26.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.kimi.com/en/help/membership/membership-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.kimi.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.kimi.com/en/help/membership/membership-pricing&lt;/a&gt; - yen pricing ¥49/¥99/¥199/¥699 and shared credit pool rules, as of 2026-09-26.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.digitalapplied.com/blog/kimi-code-k3-hands-on-setup-plans-cache-2026&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.digitalapplied.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.digitalapplied.com/blog/kimi-code-k3-hands-on-setup-plans-cache-2026&lt;/a&gt; - July 17, 2026 snapshot of listed vs promo prices and the separation banner.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://aihackers.net/value/deals/kimi-haggle/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=aihackers.net&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://aihackers.net/value/deals/kimi-haggle/&lt;/a&gt; - documented February 2026 first-month offers of $0.99 to $11.99, marked historical.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.noemititarenco.com/blog/kimi-code-plan-after-4-months-of-use-honest-review/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.noemititarenco.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.noemititarenco.com/blog/kimi-code-plan-after-4-months-of-use-honest-review/&lt;/a&gt; - skeptical four-month user review and cancellation rationale.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>MiniMax Coding Plan</title>
      <link>https://tomrochette.com/agents/model-access/minimax-coding-plan/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/minimax-coding-plan/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>coding-subscription</category><category>minimax</category><category>pricing</category>
      <description>&lt;p&gt;The product sold as the MiniMax Coding Plan became the MiniMax Token Plan on June 1, 2026, a usage-based subscription on platform.minimax.io that meters the M-series models through a Subscription Key.&#xA;&lt;strong&gt;MiniMax converted a liked flat coding plan into token billing overnight, without notice, and the resulting trust deficit is still the product&amp;rsquo;s largest liability.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;MiniMax (a Hong Kong-listed Chinese AI lab) sells Plus, Max, and Ultra tiers, with quotas shared across text, image, and speech models including M3 and M2.7.&#xA;Usage runs through a Subscription Key separate from the pay-as-you-go API key, and Anthropic-compatible endpoints let Claude Code-style harnesss plug in directly.&#xA;Quota uses a 5-hour rolling window plus a weekly window, with no carryover.&#xA;Credits packages at 1,000 per dollar cover overflow once subscription quota runs out.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Active at company scale but turbulent: H1 2026 revenue hit $117M (up 283% year over year), ARR passed $800M by August, and July token consumption ran 20 times January&amp;rsquo;s.&#xA;The stock fell more than 80% from its HK$1,330 peak to HK$298 by September 21, 2026, and JPMorgan downgraded it in June, citing M3&amp;rsquo;s lack of pricing power.&#xA;The June 2 apology and compensation package followed developer complaints after the billing switch, with a refund portal on June 3.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;One token pool covers code, images, and speech, which suits mixed-agent workloads other coding plans exclude.&lt;/li&gt;&#xA;&lt;li&gt;Credits overflow means a task does not hard-stop at the quota edge, it just starts spending Credits.&lt;/li&gt;&#xA;&lt;li&gt;Even the top tier ($132) undercuts Western $200 plans by more than a third, on an open-weight model family.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;An open GitHub issue from June 3, 2026 documents quota draining with zero API calls and an unauditable cache discount, with Plus reportedly exhausted in 4-5 hours of agent work.&lt;/li&gt;&#xA;&lt;li&gt;The FAQ states the plan does not support refunds and that MiniMax itself recommends pay-as-you-go for production.&lt;/li&gt;&#xA;&lt;li&gt;Dynamic rate limiting tightens weekday peak hours (15:00-17:30), cutting Plus to roughly 3-4 concurrent agents.&lt;/li&gt;&#xA;&lt;li&gt;Plus rose from $20 to $22 within three months of launch, an early sign that sticker prices here move fast.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;As of 2026-09-26: Plus $22/month, Max $55/month, Ultra $132/month, all with 5-hour rolling and weekly windows.&#xA;Typical peak-hour capacity is 3-4 (Plus), 4-5 (Max), and 6-7 (Ultra) agents.&#xA;Credits cost $5, $25, or $100 at 1,000 per dollar, valid 365 days, and a 10% referral discount applies at checkout.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-06-01&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Coding Plan -&amp;gt; Token Plan&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Flat per-task Coding Plan replaced by usage-based Token Plan at the M3 launch, without advance notice&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://news.aibase.com/news/28699&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=news.aibase.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://news.aibase.com/news/28699&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-06-03&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Token Plan Plus&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Observed at $20/month, $200/year&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://github.com/MiniMax-AI/MiniMax-M2.7/issues/47&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=github.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://github.com/MiniMax-AI/MiniMax-M2.7/issues/47&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Token Plan Plus/Max/Ultra&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Listed at $22/$55/$132, with Plus up $2 against June&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://platform.minimax.io/docs/guides/pricing-token-plan&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=platform.minimax.io&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://platform.minimax.io/docs/guides/pricing-token-plan&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Kimi Code (../kimi-code/index.md) sells time-windowed quota of one lab&amp;rsquo;s models and fits steady daily coders better than bursty multimodal users.&#xA;NanoGPT (../nanogpt/index.md) has no quota windows at all, removing the mid-sprint lockout risk at the cost of per-token metering.&#xA;MiniMax&amp;rsquo;s own pay-as-you-go API is the right choice for production, by the company&amp;rsquo;s own documentation.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for tinkerers who want cheap, broad token volume across modalities and will watch the usage bar.&#xA;Not for anyone needing auditable billing, refunds, or production reliability.&#xA;I will claim something arguable: until the quota denominator and per-call billing history are published, the usage bar is a trust position rather than a meter, and the June apology does not change that.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/kimi-code/&#34; &gt;Kimi Code&lt;/a&gt; - the window-quota counterpart from Moonshot, with its own 2026 pricing drama.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/nanogpt/&#34; &gt;NanoGPT&lt;/a&gt; - the no-windows gateway alternative for open-weight breadth.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/opencode/&#34; &gt;OpenCode&lt;/a&gt; - a harness commonly used to instrument quota burn on plans like this.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-selection-for-coding-tasks/&#34; &gt;Model selection for coding tasks&lt;/a&gt; - decide the model first, then plan versus PAYG.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://platform.minimax.io/docs/guides/pricing-token-plan&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=platform.minimax.io&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://platform.minimax.io/docs/guides/pricing-token-plan&lt;/a&gt; - current tiers $22/$55/$132, quota windows, agent counts, Credits packages, as of 2026-09-26.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://platform.minimax.io/docs/token-plan/faq&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=platform.minimax.io&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://platform.minimax.io/docs/token-plan/faq&lt;/a&gt; - Subscription Key mechanics, no-refund policy, production guidance, peak-hour rate limiting.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://github.com/MiniMax-AI/MiniMax-M2.7/issues/47&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=github.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://github.com/MiniMax-AI/MiniMax-M2.7/issues/47&lt;/a&gt; - critical billing-bug report: passive quota drain, unverifiable cache discount, $20 Plus in June 2026.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://news.aibase.com/news/28699&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=news.aibase.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://news.aibase.com/news/28699&lt;/a&gt; - the June 2, 2026 apology, compensation package, and Coding Plan to Token Plan switch.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://eu.36kr.com/en/p/3993817731234817&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=eu.36kr.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://eu.36kr.com/en/p/3993817731234817&lt;/a&gt; - critical business reporting: developer backlash, JPMorgan downgrade, market value collapse, H1 2026 financials.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>Model access</title>
      <link>https://tomrochette.com/agents/model-access/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/</guid>
      <category>agents</category><category>model-access</category>
      <description>&lt;p&gt;The layer that sells access to models themselves: gateways and routers metering a percentage, vendor plans selling a quota (including the consumer subscriptions that carry Claude Code and Codex), and flat subscriptions selling a ceiling.&#xA;Editors and harnesses live in their own categories; this is where the token bill gets paid.&lt;/p&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/cerebras-code/&#34; &gt;Cerebras Code&lt;/a&gt; - wafer-scale inference sold as speed, $50/$200 per month, currently sold out.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chatgpt-plans/&#34; &gt;ChatGPT plans&lt;/a&gt; - OpenAI&amp;rsquo;s subscription ladder from Go $8 to Pro 20x at $200, every tier carrying Codex.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chutes/&#34; &gt;Chutes&lt;/a&gt; - decentralized inference with pay-as-you-go plus $10/$20 subscriptions capped at 5x pay-as-you-go value.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/claude-plans/&#34; &gt;Claude plans&lt;/a&gt; - Anthropic&amp;rsquo;s Free/Pro/Max subscriptions, the only non-API way to run Claude Code, from $20 to $200 per month.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;GLM Coding Plan&lt;/a&gt; - Z.AI&amp;rsquo;s flat monthly quota for the GLM line, from $18, restructured twice since launch.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/google-ai-plans/&#34; &gt;Google AI plans&lt;/a&gt; - the Google Plus/Pro/Ultra ladder carrying Antigravity and Gemini CLI, with credits over rate limits.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/kimi-code/&#34; &gt;Kimi Code&lt;/a&gt; - Moonshot&amp;rsquo;s membership ladder for coding, $19 to $199 monthly with Code from the second tier.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/minimax-coding-plan/&#34; &gt;MiniMax Coding Plan&lt;/a&gt; - token-quota subscriptions for the MiniMax line, $22 to $132 per month, born from a silent plan replacement.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/nanogpt/&#34; &gt;NanoGPT&lt;/a&gt; - the community aggregator: hundreds of routes pay-as-you-go plus a $12 open-weight subscription.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/opencode-go/&#34; &gt;OpenCode Go&lt;/a&gt; - the OpenCode team&amp;rsquo;s $10/month open-model token pack, usable from any agent.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/opencode-zen/&#34; &gt;OpenCode Zen&lt;/a&gt; - the OpenCode team&amp;rsquo;s curated pay-per-use gateway over benchmarked endpoints.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/openrouter/&#34; &gt;OpenRouter&lt;/a&gt; - the largest model gateway, passthrough tokens plus a 5.5% credit fee, now joining Stripe.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/qwen-coding-plan/&#34; &gt;Qwen Coding Plan&lt;/a&gt; - Alibaba Cloud&amp;rsquo;s flat quota for the Qwen line plus bundled rivals, in transition to a Token Plan.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/requesty/&#34; &gt;Requesty&lt;/a&gt; - EU-residency gateway charging a flat 5% markup on upstream spend.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/supergrok/&#34; &gt;SuperGrok&lt;/a&gt; - xAI&amp;rsquo;s subscription ladder from $30 to $300, one shared weekly pool across Grok chat, Grok Build, and API.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/synthetic/&#34; &gt;Synthetic&lt;/a&gt; - a flat $30/month subscription for open-weight coding LLMs aimed at agent users.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&lt;p&gt;Its members are compared on shared rows in the &lt;a href=&#34;https://tomrochette.com/agents/model-access/model-access-feature-matrix/&#34; &gt;Model Access Feature Matrix&lt;/a&gt;.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Added Cerebras Code.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added Chutes.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added GLM Coding Plan.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added Kimi Code.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added MiniMax Coding Plan.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added NanoGPT.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added OpenCode Go.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added OpenCode Zen.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added OpenRouter.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added Requesty.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added Synthetic.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added ChatGPT plans.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added Claude plans.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added Google AI plans.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added Qwen Coding Plan.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Added SuperGrok.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>Model Access Feature Matrix</title>
      <link>https://tomrochette.com/agents/model-access/model-access-feature-matrix/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/model-access-feature-matrix/</guid>
      <category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>comparison</category><category>model-access</category><category>llm-pricing</category>
      <description>&lt;p&gt;This matrix compares the sixteen model access providers in this category, the layer that sells you tokens rather than an editor or a harness, from passthrough gateways to the consumer subscriptions that carry Claude Code, Codex, Antigravity, and Grok Build.&lt;/p&gt;&#xA;&lt;p&gt;&lt;strong&gt;The same token carries a toll from 0 to 8 percent or sits inside a flat monthly price, and the flat plans only win if your volume actually reaches their quotas.&lt;/strong&gt;&lt;/p&gt;&#xA;&lt;p&gt;Legend: ✓ yes, ✗ no, ~ partial, ? not verified.&#xA;Every cell traces to the column&amp;rsquo;s research note; volatile cells carry their own dates.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;The matrix&#xA;    &lt;div id=&#34;the-matrix&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#the-matrix&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;| Feature | &lt;a href=&#34;https://tomrochette.com/agents/model-access/cerebras-code/&#34; &gt;Cerebras Code&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/chatgpt-plans/&#34; &gt;ChatGPT plans&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/chutes/&#34; &gt;Chutes&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/claude-plans/&#34; &gt;Claude plans&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;GLM Coding Plan&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/google-ai-plans/&#34; &gt;Google AI plans&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/kimi-code/&#34; &gt;Kimi Code&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/minimax-coding-plan/&#34; &gt;MiniMax Coding Plan&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/nanogpt/&#34; &gt;NanoGPT&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/opencode-go/&#34; &gt;OpenCode Go&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/opencode-zen/&#34; &gt;OpenCode Zen&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/openrouter/&#34; &gt;OpenRouter&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/qwen-coding-plan/&#34; &gt;Qwen Coding Plan&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/requesty/&#34; &gt;Requesty&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/supergrok/&#34; &gt;SuperGrok&lt;/a&gt; | &lt;a href=&#34;https://tomrochette.com/agents/model-access/synthetic/&#34; &gt;Synthetic&lt;/a&gt; |&#xA;| &amp;mdash; | &amp;mdash; | &amp;mdash; | &amp;mdash; | &amp;mdash; | &amp;mdash; | &amp;mdash; | &amp;mdash; | &amp;mdash; | &amp;mdash; | &amp;mdash; | &amp;mdash; |&#xA;| Kind | speed-first inference subscription | vendor consumer subscription (OpenAI) | decentralized inference market with subscriptions | vendor consumer subscription (Anthropic) | vendor coding plan (Z.AI) | vendor consumer subscription (Google) | vendor membership ladder (Moonshot) | vendor token plan | community aggregator plus subscription | token-pack subscription | curated pay-per-use gateway | passthrough gateway | vendor coding plan (Alibaba Cloud) | governance gateway | vendor consumer subscription (xAI) | flat open-model subscription |&#xA;| Billing mechanics | flat monthly with token-per-day caps | flat monthly, credit-based usage in 5-hour windows over a shared pool | PAYG plus $10/$20 tiers adding quota and 6-10% off | flat monthly, usage multipliers over a shared 5-hour and weekly clock | flat monthly converted to credits, peak multipliers | flat monthly over rate limits plus a flexible AI credit pool | monthly membership tiers over a shared quota pool | monthly token quotas in 5-hour and weekly windows | PAYG deposits plus a $12 open-weight subscription | $10/month for per-model dollar bundles | per-token, no markup claimed, card fees at cost | passthrough tokens, 5.5% credit fee ($0.80 minimum, Business 8%) | flat monthly converted to request or credit quotas with 5-hour, weekly, and monthly caps | flat 5% of upstream spend | flat monthly over one shared weekly pool spanning chat, Build, and API | $30 pack or per-token |&#xA;| Cheapest paid entry | Pro $50/month (sold out 2026-09-26) | Go $8/month | Plus $10/month (PAYG from cents) | Pro $20/month ($17 annual) | Lite $18/month | AI Plus (about CA$13.99/month on the fetched Canadian page) | Moderato $19/month ($15 annual) | Plus $22/month | Pro $12/month (PAYG from $0.10) | $10/month | $20 auto-reload, usage-priced | free tier, then the credit fee | Token Plan Lite $6/month (list $8, Singapore) | free tier, then 5% | SuperGrok $30/month (Lite $10 in testing) | $30/month pack |&#xA;| Top tier | Max $200/month | Pro 20x $200/month | Pro $20/month | Max 20x $200/month | Max $168/month | AI Ultra 20x (about $200/month, secondhand) | Vivace $199/month ($159 annual) | Ultra $132/month | the $12 Pro tier is the top | one $10/month tier | none, usage-based | Enterprise custom | Token Plan Pro $68/month (list $80) | Enterprise custom | SuperGrok Heavy $300/month (third-party reported) | stacked $30 packs |&#xA;| Models you can reach | fast Qwen and GLM coding lines | the GPT-6 and GPT-5.6 lines | open weights (GLM-5.2, Kimi K3, Qwen, DeepSeek) | the Claude line, Fable via credits on Pro | the GLM line | the Gemini line via Antigravity and Gemini CLI | the Kimi line, K3 from Pro up | the MiniMax line | hundreds of routes, open weights included | open coding models only (Qwen, Grok, Luna lines) | open plus Claude, GPT, Gemini, and Grok lines | the largest catalog, frontier and open | the Qwen line plus bundled Kimi, GLM, and MiniMax on some tiers | 600+ models | the Grok line (4.6, 4.7) via Grok Build | the open-weight always-on lineup |&#xA;| Limit or quota form | 24M/120M tokens per day, TPM caps | shared agentic pool, published per-model message ranges per 5 hours | daily quota, then discounted PAYG, capped at 5x value | one shared pool, unpublished multipliers (1x/5x/20x), 5-hour plus weekly clocks | credits per 5 hours and week, 3x peak burn on flagships | 5-hour Ultra credit refresh, Never or Always overage modes | one shared pool across chat, research, and code | 5-hour and weekly windows, dynamic peak limits | 60M input tokens per week | per-model monthly dollars ($15-$60), 5-hour and weekly windows | balance only, auto-reload below $5 | provider rate limits pass through | whichever of the 5-hour, weekly, or monthly caps hits first | budget caps you set | one weekly pool shared across chat, Build, and API | 500 requests per 5 hours per pack, 1 concurrent per model |&#xA;| The catch | sold out, and measured speed fell far short of the marketing | Work, images, and voice drain the same pool; real cost lands near $100-$200 per developer once heavy | terms changed three times in early 2026, capacity rides on miners | chat and code share one pool, quotas unpublished, Fable capped at 50% of weekly limits on Max | non-refundable, peak burn cuts effective quota, China-based routing | US dollar prices render client-side, credit purchase needs sign-in, storage numbers disagree across Google pages | quota shared with chat and Deep Research, ladder moved twice in 2026 | quota drained with zero calls per a June 2026 issue, no refunds | the deal narrowed from $8 unlimited to $12 capped, opaque governance | no longer general API access, agent traffic only, monthly catalog churn | measured at more than 4x OpenRouter on identical open models | the fee punishes small top-ups, credits can expire within a year | Coding Plan Pro sold out with no migration path, and the predecessor plan suspended automation users | the 5% grows with spend, hosted-only, 30-day default log retention | Lite and Heavy dollar figures are not on the official card, and the shared pool means chat competes with your agent | the 3x-Claude framing is contested, packs stack for parallelism, pinned models eventually 404 |&#xA;| Frontier models | ✗ open lines only | ✓ own frontier | ✗ open weights | ✓ own frontier | ✗ GLM only | ✓ own frontier | ✗ Kimi only | ✗ MiniMax only | ~ closed models on PAYG, open in the sub | ✗ open models only | ✓ Claude, GPT, Gemini, Grok | ✓ nearly all | ✓ own frontier | ✓ 600+ including frontier | ✓ own frontier | ✗ open models only |&#xA;| Works from any agent | ✓ API key | ~ subscription locked to first-party surfaces, the CLI rides an API key instead | ✓ OpenAI-compatible | ✗ subscription locked to first-party harnesses, third-party use rides API keys | ~ supported tools only, special endpoints | ~ Antigravity and Gemini CLI first-party, other agents ride API keys | ~ first-party clients emphasized, API on higher tiers | ✓ OpenAI-compatible | ✓ OpenAI-compatible | ✓ any agent, coding-agent headers required since 2026-09 | ✓ OpenAI/Anthropic/Google-compatible endpoints | ✓ one key, any client | ~ one key across coding tools, automation restricted | ✓ one OpenAI-compatible endpoint | ~ Grok Build first-party, partner sign-in for others, API key otherwise | ✓ OpenAI-compatible |&#xA;| Price trajectory | stable $50/$200 since 2025-08 | Pro split into $100/$200, Business replaced Team, credit billing since April | free tier retired, Base tier cut | prices stable since the Max launch, but the quota behind them moved twice in spring 2026 | $6/$30 to $18/$80/$168 within a year | AI Premium renamed Plus and the ladder gained a Plus rung | promos became standard pricing 2026-09 | Plus $20 to $22 in three months, plan replaced overnight | $8 to $12 with a new cap | $5 promo removed, $15 flagship standard | baseline 2026-09-26, heavy deprecation churn | fee flattened to 5.5% (2025-06), Business 8% added (2026-09) | Lite discontinued, Pro sold out, the Token Plan restructure replaced both | unchanged since tracking began (2026-09-26) | Lite announced in March, Plus observed by September | $20/$60 tiers replaced by a $30 pack |&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Reading the matrix&#xA;    &lt;div id=&#34;reading-the-matrix&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#reading-the-matrix&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;I read the billing-mechanics row first, because it decides who this provider is for: passthrough gateways meter reality, vendor plans sell a quota, and flat subscriptions sell a ceiling.&#xA;&lt;strong&gt;The gateways (OpenRouter, Requesty, Zen) monetize a percentage, the vendor plans (GLM, Kimi, MiniMax, Qwen, and the consumer subscriptions) monetize commitment, and the community subs (NanoGPT, Synthetic, Chutes, Go) monetize the gap between list price and what self-hosted open models actually cost to serve.&lt;/strong&gt;&#xA;On price trajectories, only Requesty, Zen, and the Claude price tags have not moved, and Zen&amp;rsquo;s baseline is one day old, so the stable-looking rows are the youngest ones.&#xA;The frontier-models row is the real segmentation: if your loop needs Claude, GPT, Gemini, or Grok, the flat open-model subscriptions are irrelevant and the choice is gateway versus vendor subscription.&#xA;The catches cluster into two kinds: mechanical limits you can engineer around (concurrency, windows, peak multipliers) and trust problems you cannot (quota draining, silent plan replacement, narrowed deals), and I weight the second kind higher.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Choosing an access provider&#xA;    &lt;div id=&#34;choosing-an-access-provider&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#choosing-an-access-provider&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Default BYOK loop across many providers: OpenRouter for breadth, Requesty if EU residency or budget governance is required.&lt;/li&gt;&#xA;&lt;li&gt;First-party subscription loops: Claude plans, ChatGPT plans, Google AI plans, or SuperGrok, matched to the model family your daily driver uses.&lt;/li&gt;&#xA;&lt;li&gt;One vendor&amp;rsquo;s models all day: that vendor&amp;rsquo;s plan (GLM, Kimi, MiniMax, Qwen), after checking the peak-hour mechanics against your working hours.&lt;/li&gt;&#xA;&lt;li&gt;Open models at a flat ceiling: OpenCode Go or Synthetic for agent-shaped workloads, NanoGPT for breadth beyond coding, Chutes for the cheapest per-token open rates.&lt;/li&gt;&#xA;&lt;li&gt;Frontier models without a subscription: OpenCode Zen, paying the measured curation premium only where a mis-served provider would silently degrade your agent.&lt;/li&gt;&#xA;&lt;li&gt;Raw speed: Cerebras Code, when it is in stock and your context fits the window.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created with eleven columns when the owner-directed Model access category was seeded.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Extended from eleven to thirteen columns with ChatGPT plans and Claude plans on owner instruction, re-sorted; the vendor consumer subscriptions that carry Claude Code and Codex joined the category.&lt;/li&gt;&#xA;&lt;li&gt;2026-09-26 - Extended from thirteen to sixteen columns with Qwen Coding Plan, Google AI plans, and SuperGrok on the owner&amp;rsquo;s completeness question, re-sorted; every frontier vendor&amp;rsquo;s subscription now has a column.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model Provider Feature Matrix&lt;/a&gt; - the vendors behind these access products, compared as bundles&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-selection-for-coding-tasks/&#34; &gt;Model Selection for Coding Tasks&lt;/a&gt; - the per-model, per-token economics this layer packages&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/opencode/&#34; &gt;OpenCode&lt;/a&gt; - the harness whose Zen gateway and Go subscription anchor two columns here&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/trackers-and-leaderboards/trackers-and-leaderboards-feature-matrix/&#34; &gt;Trackers and Leaderboards Feature Matrix&lt;/a&gt; - where usage and price signals about this layer get measured&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://opencode.ai/docs/go&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=opencode.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://opencode.ai/docs/go&lt;/a&gt; - Go subscription tiers, model limit table (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://opencode.ai/docs/zen/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=opencode.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://opencode.ai/docs/zen/&lt;/a&gt; - Zen per-model price table and free models (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://openrouter.ai/docs/faq&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=openrouter.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://openrouter.ai/docs/faq&lt;/a&gt; - OpenRouter billing and routing FAQ (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://requesty.ai/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=requesty.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://requesty.ai/pricing&lt;/a&gt; - Requesty tier table and 5% markup quote (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://nano-gpt.com/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=nano-gpt.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://nano-gpt.com/pricing&lt;/a&gt; - NanoGPT subscription and per-token prices (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://platform.minimax.io/docs/guides/pricing-token-plan&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=platform.minimax.io&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://platform.minimax.io/docs/guides/pricing-token-plan&lt;/a&gt; - MiniMax Token Plan tiers (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://docs.z.ai/devpack/overview&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=docs.z.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://docs.z.ai/devpack/overview&lt;/a&gt; - GLM Coding Plan documentation (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://chutes.ai/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=chutes.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://chutes.ai/pricing&lt;/a&gt; - Chutes PAYG and subscription tiers (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://synthetic.new/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=synthetic.new&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://synthetic.new/pricing&lt;/a&gt; - Synthetic subscription pack pricing (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>NanoGPT</title>
      <link>https://tomrochette.com/agents/model-access/nanogpt/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/nanogpt/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>model-gateway</category><category>open-weights</category><category>privacy</category>
      <description>&lt;p&gt;NanoGPT (nano-gpt.com) is an independent model-access gateway combining pay-per-token API access at list prices with a flat subscription that bundles open-weight model usage at no per-token cost.&#xA;&lt;strong&gt;At $12 per month it is the cheapest flat-rate open-weight tier I can verify anywhere, but the deal quietly shrank over the past year.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;The service runs a hosted web chat and an OpenAI-compatible API at provider list prices with no markup, taking deposits from $0.10 in crypto or $1 by card.&#xA;The Pro subscription includes open-weight text and image models, currently framed as 60 million input tokens per week plus a 5% discount on eligible paid text models.&#xA;Coverage spans text, image, video, 3D, audio, and embedding models under one balance.&#xA;An independent operator runs it, and the pages I fetched name no corporate parent.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Active: the blog published within the week of verification (September 22, 2026), ships monthly crypto-payment statistics, and added zero-data-retention routing on September 3, 2026.&#xA;The Wayback Machine holds 57 captures from September 2024 through September 2026, so the service has operated continuously for at least two years.&#xA;Its own Hacker News footprint is nearly nil: two stories in late 2024 with zero comments, while the NanoGPT name on HN belongs to Karpathy&amp;rsquo;s training repo (a 1,532-point story).&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;List-price API with the exact cost printed on every request makes it a rare auditable gateway.&lt;/li&gt;&#xA;&lt;li&gt;The subscription undercuts every first-party lab plan while covering text and image generation.&lt;/li&gt;&#xA;&lt;li&gt;The privacy posture is concrete: crypto payments (Monero led deposits at 38.55% in August 2026), no-prompt-logging claims, and ZDR routing.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;The deal narrowed: September 2025 sold unlimited personal open-weight use at $8, September 2026 sells 60M input tokens per week at $12, a 50% rise with a new cap.&lt;/li&gt;&#xA;&lt;li&gt;Bring-your-own-key requests are billed at 5% of normal model cost, and pinning a specific provider adds another 5%.&lt;/li&gt;&#xA;&lt;li&gt;Governance is opaque: no funding, ownership, or company facts appear on fetched pages, so nobody is accountable if terms move again.&lt;/li&gt;&#xA;&lt;li&gt;Scrutiny lives in Discord and niche communities rather than HN, so fewer independent eyes check the claims.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Pro is $12/month as of 2026-09-26, including 60M input tokens per week, the 5% paid-text-model discount, and web plus API access.&#xA;Pay-as-you-go needs no subscription: a free tier with one web-only model, list-price API billing, and $0.10 crypto or $1 card deposit minimums.&#xA;In September 2025, Pro was $8/month with unlimited personal open-weight usage capped at 60,000 generations per month and 2,000 per day.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2025-09-01&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pro&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;$8/month, unlimited personal open-weight usage (60,000 generations/month, 2,000/day)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;http://web.archive.org/web/20250901063827/https://nano-gpt.com/subscription&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=web.archive.org&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;http://web.archive.org/web/20250901063827/https://nano-gpt.com/subscription&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pro&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;$12/month, 60M included input tokens per week, 5% paid-text-model discount&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://nano-gpt.com/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=nano-gpt.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://nano-gpt.com/pricing&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Kimi Code (../kimi-code/index.md) is the pick when you want one lab&amp;rsquo;s frontier open-weight models with quota predictability at a similar monthly price.&#xA;MiniMax Coding Plan (../minimax-coding-plan/index.md) covers multimodal token volume cheaply but locks you out in 5-hour windows, which NanoGPT never does.&#xA;The free PAYG path also makes NanoGPT the lowest-commitment way to test a model before any subscription.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for privacy-minded tinkerers and open-weight power users who want flat-cost breadth without quota-window lockouts.&#xA;Not for engineers whose work needs closed frontier models at subscription quotas, since the $12 tier&amp;rsquo;s value lives almost entirely in open weights.&#xA;I will claim something arguable: at $12 with open weights included, this beats any $15-$20 first-party plan on value, because open-weight quality has closed enough of the gap that the labs&amp;rsquo; premium is now a convenience fee.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/kimi-code/&#34; &gt;Kimi Code&lt;/a&gt; - the first-party open-weight subscription to compare quota mechanics against.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/minimax-coding-plan/&#34; &gt;MiniMax Coding Plan&lt;/a&gt; - the token-metered alternative with modality breadth.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model provider feature matrix&lt;/a&gt; - where gateways sit against first-party plans.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-selection-for-coding-tasks/&#34; &gt;Model selection for coding tasks&lt;/a&gt; - choose models on evidence before choosing where to buy them.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://nano-gpt.com/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=nano-gpt.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://nano-gpt.com/pricing&lt;/a&gt; - current Pro pricing ($12/month, 60M input tokens/week), PAYG minimums, 5% BYOK and pinning fees, as of 2026-09-26.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;http://web.archive.org/web/20250901063827/https://nano-gpt.com/subscription&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=web.archive.org&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;http://web.archive.org/web/20250901063827/https://nano-gpt.com/subscription&lt;/a&gt; - September 2025 capture showing $8/month with unlimited personal open-weight usage.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://nano-gpt.com/blog&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=nano-gpt.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://nano-gpt.com/blog&lt;/a&gt; - activity evidence: September 2026 posts, ZDR routing, monthly crypto-payment statistics.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://hn.algolia.com/api/v1/search?query=nano-gpt.com&amp;amp;restrictSearchableAttributes=url&amp;amp;tags=story&amp;amp;hitsPerPage=10&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=hn.algolia.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://hn.algolia.com/api/v1/search?query=nano-gpt.com&amp;restrictSearchableAttributes=url&amp;tags=story&amp;hitsPerPage=10&lt;/a&gt; - the service&amp;rsquo;s thin HN footprint (two 0-comment stories, 2024).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://hn.algolia.com/api/v1/search?query=NanoGPT&amp;amp;tags=story&amp;amp;hitsPerPage=10&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=hn.algolia.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://hn.algolia.com/api/v1/search?query=NanoGPT&amp;tags=story&amp;hitsPerPage=10&lt;/a&gt; - the name collision with Karpathy&amp;rsquo;s repo dominating HN results.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>OpenCode Go</title>
      <link>https://tomrochette.com/agents/model-access/opencode-go/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/opencode-go/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>coding-subscription</category><category>open-models</category>
      <description>&lt;p&gt;OpenCode Go is a $10/month subscription from the OpenCode (Anomaly) team that bundles access to a curated set of open coding models, usable from OpenCode or any compatible coding agent.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Go sells model access, not an editor or a harness.&#xA;You subscribe, copy an API key, and point any agent at OpenAI-compatible, Anthropic-compatible, or Responses endpoints under &lt;code&gt;opencode.ai/zen/go/v1&lt;/code&gt;.&#xA;The lineup is 32 open-weight coding models as of 2026-09-26 (Grok 4.7/4.6, GLM-5.3 family, Kimi K3, Qwen3.x, DeepSeek V4, MiniMax M3, MiMo, GPT 6 Luna), and only one member per workspace can hold a subscription.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Active and changing fast.&#xA;The OpenCode repo shows about 208K GitHub stars as of 2026-09-26, the Go docs were last updated 2026-09-25, and the plan has grown from a three-model Beta in March 2026 to 32 models.&#xA;Churn is constant: GPT 6 Luna arrived 2026-09-22 and Space Bunny Free appeared as a limited-time unlimited model.&#xA;&lt;strong&gt;Go converts $10 into up to $60 of metered usage per month, and that 6x ratio is the entire value proposition.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;The $60 monthly usage value costs $10, up to 6x leverage if you max it.&lt;/li&gt;&#xA;&lt;li&gt;Models are benchmarked for agentic coding before inclusion, per the team.&lt;/li&gt;&#xA;&lt;li&gt;Works with any agent; Claude Code, Codex, Pi, jcode, and Kilo Code CLI are validated clients.&lt;/li&gt;&#xA;&lt;li&gt;Dollar-denominated windows make the bill predictable, free models stay available after limits, and &amp;ldquo;Use balance&amp;rdquo; falls back to your Zen balance instead of hard-stopping requests.&lt;/li&gt;&#xA;&lt;li&gt;Most models run with zero retention and no training on your data.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;strong&gt;This is no longer general API access.&lt;/strong&gt;&lt;/li&gt;&#xA;&lt;li&gt;Since September 2026, clients must send coding-agent traffic, a real user agent, and an &lt;code&gt;x-opencode-session&lt;/code&gt; header, which the community reads as whitelisting non-agent use.&lt;/li&gt;&#xA;&lt;li&gt;Limits are dollar-value, so flagship tiers (Grok, Kimi K3, GLM-5.3, GPT 6 Luna) include only $15/month, with a $3 5-hour window.&lt;/li&gt;&#xA;&lt;li&gt;The catalog and limits change monthly, the $5 first-month promo is gone, and usage estimates assume heavy prompt caching, so your real request counts will not match.&lt;/li&gt;&#xA;&lt;li&gt;Muse Spark Contributor models train on your prompts and are region-limited.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;$10/month, cancel any time, as of 2026-09-26.&#xA;Each model carries a monthly dollar limit (mostly $60, some $30, $15 for flagships), with windows at 20% per rolling 5 hours (so $12 on a $60 model) and 50% weekly ($30).&#xA;Exceeded limits fall back to free models, or to your Zen balance if you enable it.&#xA;Top-ups draw on the shared Zen balance, where card fees are passed at cost (4.4% + $0.30).&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-03-12&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Go (Beta)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: $10/month, three models (GLM-5, Kimi K2.5, MiniMax M2.5), $60 usage value&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://help.apiyi.com/en/opencode-go-subscription-worth-it-review-en.html&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=help.apiyi.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://help.apiyi.com/en/opencode-go-subscription-worth-it-review-en.html&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-08-24&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Go&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;First-month $5 promo removed, flat $10/month&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://codingplan.org/en/plans/opencode-go&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=codingplan.org&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://codingplan.org/en/plans/opencode-go&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-02&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Go&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;$15 monthly limit became the standard for several flagship models (community-reported)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://www.reddit.com/r/opencode/comments/1vo9j8l/opencode_go_15_is_now_the_standard/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.reddit.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.reddit.com/r/opencode/comments/1vo9j8l/opencode_go_15_is_now_the_standard/&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-24&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;GLM-5.3-Flash&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Monthly limit doubled to $60 (docs confirm $60 by 2026-09-25)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://www.reddit.com/r/opencode/comments/1wc1cwe/glm_53_flash_now_gets_twice_the_limits_on/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.reddit.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.reddit.com/r/opencode/comments/1wc1cwe/glm_53_flash_now_gets_twice_the_limits_on/&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;OpenCode Zen (../opencode-zen/index.md) is the sibling pay-per-use gateway: no subscription, huge catalog including Claude and GPT, but per-token prices that can exceed OpenRouter&amp;rsquo;s.&#xA;Choose Go when you want a capped, predictable bill for open models.&#xA;OpenRouter (../openrouter/index.md) has hundreds of models with routing and fallback at passthrough prices plus a 5.5% credit fee; choose it when you need frontier or long-tail models Go does not carry.&#xA;Direct DeepSeek API beats Go unless you fully use the $60 cap, per a community breakdown that pegs Go at 33% cheaper only at full usage.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for heavy users of open coding models who will actually burn the $60 monthly value.&#xA;Not for anyone needing general-purpose API access, closed frontier models, or light usage.&#xA;My disagreeable claim: at roughly 50% usage Go is about break-even with direct API pricing, so most subscribers are partly paying for slack they never consume.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/opencode-zen/&#34; &gt;OpenCode Zen&lt;/a&gt; - the sibling pay-per-use gateway sharing the same balance and free models.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/openrouter/&#34; &gt;OpenRouter&lt;/a&gt; - the largest general gateway and the main price competitor for the same open models.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/opencode/&#34; &gt;OpenCode&lt;/a&gt; - the harness Go is bundled with and defaults to.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model provider feature matrix&lt;/a&gt; - cross-provider comparison of access features.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://opencode.ai/go&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=opencode.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://opencode.ai/go&lt;/a&gt; - product page, $10/month framing, model snapshot, GitHub star count (200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://opencode.ai/docs/go&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=opencode.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://opencode.ai/docs/go&lt;/a&gt; - full model/limit table, 20%/50%/100% windows, validated clients, header requirements (200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://codingplan.org/en/plans/opencode-go&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=codingplan.org&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://codingplan.org/en/plans/opencode-go&lt;/a&gt; - $5 first-month promo removal on 2026-08-24, catalog churn timeline (200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://help.apiyi.com/en/opencode-go-subscription-worth-it-review-en.html&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=help.apiyi.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://help.apiyi.com/en/opencode-go-subscription-worth-it-review-en.html&lt;/a&gt; - March 2026 Beta state, dollar-window mechanics, critical take (200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.reddit.com/r/opencode/comments/1w9pyvq/opencode_go_is_no_longer_general_api_access/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.reddit.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.reddit.com/r/opencode/comments/1w9pyvq/opencode_go_is_no_longer_general_api_access/&lt;/a&gt; - community reaction to session-header requirement (direct fetch 403; content read via search).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.reddit.com/r/opencodeCLI/comments/1tqx9u1/why_opencode_gos_deepseek_v4_pro_is_33_cheaper/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.reddit.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.reddit.com/r/opencodeCLI/comments/1tqx9u1/why_opencode_gos_deepseek_v4_pro_is_33_cheaper/&lt;/a&gt; - full-usage price math vs DeepSeek API (content read via search).&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>OpenCode Zen</title>
      <link>https://tomrochette.com/agents/model-access/opencode-zen/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/opencode-zen/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>ai-gateway</category><category>pay-per-use</category>
      <description>&lt;p&gt;OpenCode Zen is the OpenCode (Anomaly) team&amp;rsquo;s curated pay-per-use AI gateway, one API key over a tested catalog that spans open models and the Claude, GPT, Gemini, and Grok lines.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Zen is a billing and routing gateway, not a model: you add credits, get a key, and call OpenAI-compatible, Anthropic-compatible, Google, or Responses endpoints under &lt;code&gt;opencode.ai/zen/v1&lt;/code&gt;.&#xA;The team benchmarks each model/provider pair and serves what passes, their answer to getting a degraded version of a model through a generic router.&#xA;It also sells to teams (workspaces, roles, member spending caps, model toggles, BYOK for OpenAI and Anthropic keys), and free stealth and promo models rotate through the catalog.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Active, with heavy catalog churn managed through a public deprecation table (docs updated 2026-09-25), and workspaces free during the beta with team pricing unannounced.&#xA;The Reddit thread &amp;ldquo;Opencode Zen is astoundingly more expensive than OpenRouter&amp;rdquo; (July 2026) is the sharpest public criticism and remains the note&amp;rsquo;s key stress test.&#xA;&lt;strong&gt;Zen sells curation and reliability, and community measurements say that premium can exceed 4x the cheapest gateway on identical models.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Every endpoint is benchmarked for coding-agent use, so you avoid badly served providers.&lt;/li&gt;&#xA;&lt;li&gt;A stable of free models (Big Pickle, Space Bunny Free, MiMo Free, Nemotron Free) with documented data caveats.&lt;/li&gt;&#xA;&lt;li&gt;Markups are claimed to be zero beyond pass-through processing fees (4.4% + $0.30 per card transaction).&lt;/li&gt;&#xA;&lt;li&gt;US-hosted with zero-retention on most models, exceptions listed per model.&lt;/li&gt;&#xA;&lt;li&gt;Same balance backs the Go subscription&amp;rsquo;s overflow, and the key works from any agent.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;A July 2026 community test measured Zen at more than 4x OpenRouter on Kimi K2.6 and MiniMax M2.7.&lt;/li&gt;&#xA;&lt;li&gt;The gap persists per model as of 2026-09-26: Zen lists GLM-5.3-Flash at $0.15/$0.50 per 1M tokens while OpenRouter&amp;rsquo;s catalog shows about $0.05/$0.14.&lt;/li&gt;&#xA;&lt;li&gt;Auto-reload adds $20 whenever your balance drops below $5, which can overrun a monthly budget you set.&lt;/li&gt;&#xA;&lt;li&gt;Deprecated models vanish on published dates, so configs need migrations, and free stealth models may use your data to improve the model during their free window.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Pay-as-you-go per 1M tokens, as of 2026-09-26: GLM-5.3 $1.40/$4.40, GLM-5.3-Flash $0.15/$0.50, Kimi K3 $3/$15, Qwen3.7 Plus $0.40/$1.60, DeepSeek V4.1 Flash $0.30/$1.20, MiniMax M3 $0.30/$1.20.&#xA;Frontier lines: Claude Sonnet 5 $2/$10, Claude Opus 5.5 $4/$20, GPT 5.5 $5/$30, Gemini 3.8 Flash $1.50/$7.50, Grok 4.7 $2/$6.&#xA;Jev 1.13 charges $0.042 per input with free output, and nine models are free including the Big Pickle and Space Bunny stealth models, all for a limited time.&#xA;Auto-reload charges $20 when the balance falls below $5, and OpenCode itself uses low-cost models to generate session titles on your bill.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pay-per-use catalog&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: per-model rates as published (GLM-5.3 $1.40/$4.40, Kimi K3 $3/$15, Jev 1.13 $0.042 in)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://opencode.ai/docs/zen/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=opencode.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://opencode.ai/docs/zen/&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;OpenCode Go (../opencode-go/index.md) is the sibling subscription: $10/month for capped open-model usage, no frontier models, same balance; choose it when you want a ceiling.&#xA;OpenRouter (../openrouter/index.md) is the breadth play: hundreds of models, passthrough pricing, 5.5% credit fee; choose it when price per token on open models matters more than curation.&#xA;Holding your own provider keys is cheapest for a single provider, and Zen&amp;rsquo;s BYOK only covers OpenAI and Anthropic.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for OpenCode-centric developers who want vetted endpoints, free models to fall back on, and US-hosted zero-retention.&#xA;Not for price-sensitive, high-volume runs on open models, where OpenRouter wins on raw per-token cost.&#xA;My disagreeable claim: the curation premium is worth paying only when a mis-served provider would silently degrade your agent, and most coding work is not that sensitive.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/opencode-go/&#34; &gt;OpenCode Go&lt;/a&gt; - the sibling subscription that overflows into this balance.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/openrouter/&#34; &gt;OpenRouter&lt;/a&gt; - the broader, usually cheaper gateway this one is measured against.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/opencode/&#34; &gt;OpenCode&lt;/a&gt; - the harness Zen is the default recommended provider for.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model provider feature matrix&lt;/a&gt; - where Zen sits among access providers.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://opencode.ai/docs/zen/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=opencode.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://opencode.ai/docs/zen/&lt;/a&gt; - per-model price table, free models, auto-reload, deprecations, privacy (200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://docs.docker.com/ai/docker-agent/providers/opencode-zen/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=docs.docker.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://docs.docker.com/ai/docker-agent/providers/opencode-zen/&lt;/a&gt; - third-party integration doc, Zen vs Go billing table (200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.reddit.com/r/opencodeCLI/comments/1syog9v/opencode_zen_is_astoundingly_more_expensive_than/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.reddit.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.reddit.com/r/opencodeCLI/comments/1syog9v/opencode_zen_is_astoundingly_more_expensive_than/&lt;/a&gt; - critical 4x price claim vs OpenRouter (direct fetch 403; content read via search).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://costgoat.com/pricing/openrouter/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=costgoat.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://costgoat.com/pricing/openrouter/&lt;/a&gt; - OpenRouter per-model prices used for the GLM-5.3-Flash comparison (200, as of 2026-09-25).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://opencode.ai/docs/go&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=opencode.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://opencode.ai/docs/go&lt;/a&gt; - shared balance and &amp;ldquo;Use balance&amp;rdquo; fallback mechanics (200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.truefoundry.com/blog/openrouter-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.truefoundry.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.truefoundry.com/blog/openrouter-pricing&lt;/a&gt; - OpenRouter fee structure used as the price counterpoint (200).&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>OpenRouter</title>
      <link>https://tomrochette.com/agents/model-access/openrouter/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/openrouter/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>llm-gateway</category><category>pay-per-token</category>
      <description>&lt;p&gt;OpenRouter is the largest multi-provider LLM gateway, one API key and one prepaid credit balance across hundreds of models from dozens of providers, with automatic routing and fallback.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;A marketplace gateway founded in early 2023 by Alex Atallah (OpenSea co-founder).&#xA;Provider token prices pass through unchanged; you fund prepaid credits and OpenRouter takes a fee on the purchase, not on inference.&#xA;Routing variants (:nitro for speed, :floor for price, :exacto for tool-calling quality, :free and :batch) change how requests are placed, and BYOK keeps your own provider keys behind OpenRouter&amp;rsquo;s routing, analytics, and fallbacks.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;&lt;strong&gt;The scale is venture-grade: $113M Series B led by Alphabet&amp;rsquo;s CapitalG at about $1.3B post-money in May 2026, with 8M users and roughly 100 trillion tokens per month.&lt;/strong&gt;&#xA;Annualized inference spend through the platform grew from $10M (October 2024) to over $100M (May 2025), per a tracked pricing blueprint.&#xA;On 2026-08-19 OpenRouter announced it is joining Stripe, with closing expected within weeks of the announcement; the post commits to the same product, name, roadmap, and provider-neutral routing, and states the platform now processes 10+ trillion tokens per day across 400+ models for more than 10 million developers.&#xA;The pricing page was rebuilt on 2026-09-20 into four tiers (Free, Standard, Business, Enterprise), and the pricing page itself renders client-side, so my fetch returned navigation only.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Passthrough pricing is real: verified spot checks (Claude Opus 5 at $5/$25) match provider list prices.&lt;/li&gt;&#xA;&lt;li&gt;One key, one balance, automatic provider fallback, per-model price/latency comparison, and a fee structure published in full.&lt;/li&gt;&#xA;&lt;li&gt;Free tier is a genuine on-ramp: 25+ free models, 1,000 requests/day after a one-time $10 credit purchase.&lt;/li&gt;&#xA;&lt;li&gt;BYOK is free through $25,000/month of list-price inference, which covers most teams entirely.&lt;/li&gt;&#xA;&lt;li&gt;The Business tier makes EU-only or US-only routing self-serve instead of an Enterprise contract.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;strong&gt;The 5.5% credit fee with a $0.80 minimum punishes small top-ups: a $5 purchase costs 16% in fees.&lt;/strong&gt;&lt;/li&gt;&#xA;&lt;li&gt;BYOK&amp;rsquo;s free allowance is metered at OpenRouter list price, not your negotiated rate, so discounts do not slow the meter.&lt;/li&gt;&#xA;&lt;li&gt;Paid-tier rate limits are passthrough from providers, and 429s arrive without queueing or backoff; your client owns retries.&lt;/li&gt;&#xA;&lt;li&gt;No public SLA below Enterprise.&lt;/li&gt;&#xA;&lt;li&gt;Routing can silently move you to a different provider, and a provider price change flows straight to your bill.&lt;/li&gt;&#xA;&lt;li&gt;Credits may expire after one year per the terms.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Free: $0, 25+ free models, 20 requests/minute, 50 requests/day (1,000/day with $10 lifetime credits), workspace limit 5.&#xA;Standard: 5.5% fee per credit purchase, $0.80 minimum, crypto 5.0% flat, no subscription, workspace limit 5.&#xA;Business: 8% fee, inference locked to EU or US providers with no cross-region fallback, workspace limit 1,000 (launched 2026-09-07).&#xA;Enterprise: custom, volume commitments, SSO/SAML, contractual SLAs, $200,000/month free BYOK allowance.&#xA;BYOK: 5% of equivalent cost above the free allowance, as of 2026-09-26.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2023-05&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Launch&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: passthrough tokens, prepaid credits, purchase fee&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://www.usagepricing.com/blueprint/openrouter&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.usagepricing.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.usagepricing.com/blueprint/openrouter&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2025-06-09&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Credit fee&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Flattened to 5.5% with $0.80 minimum; crypto to 5.0% flat&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://www.usagepricing.com/blueprint/openrouter&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.usagepricing.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.usagepricing.com/blueprint/openrouter&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-07-14&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;BYOK&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Free allowance re-based from 1M/5M requests to $25,000/$200,000 of list-price inference&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://www.usagepricing.com/blueprint/openrouter&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.usagepricing.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.usagepricing.com/blueprint/openrouter&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-07&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Business&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;New self-serve tier at 8% fee for EU/US-only routing&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://www.usagepricing.com/blueprint/openrouter&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.usagepricing.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.usagepricing.com/blueprint/openrouter&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-20&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Table&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pay-as-you-go renamed Standard; workspace limits published (5/5/1000/Custom)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://www.usagepricing.com/blueprint/openrouter&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.usagepricing.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.usagepricing.com/blueprint/openrouter&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;OpenCode Go (../opencode-go/index.md) is a $10 subscription for open coding models with hard caps; choose OpenRouter when your usage is spiky or you need frontier models.&#xA;OpenCode Zen (../opencode-zen/index.md) is the curated coding gateway, usually pricier per token on open models; choose Zen when benchmarked endpoints and free stealth models matter more than unit cost.&#xA;Direct provider accounts are cheapest for one dominant model at scale, at the cost of N billing relationships and no cross-provider fallback.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for multi-model teams and tinkerers who value breadth, fallback, and one bill.&#xA;Not for single-provider, high-volume workloads with negotiated rates, where the fee is pure overhead.&#xA;My disagreeable claim: I would pay the 5.5% rather than run the same multi-provider setup myself, because the fee buys uptime pooling, not just convenience.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/opencode-go/&#34; &gt;OpenCode Go&lt;/a&gt; - the flat-fee subscription alternative for open coding models.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/opencode-zen/&#34; &gt;OpenCode Zen&lt;/a&gt; - the curated gateway that community benchmarks price against this one.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/trackers-and-leaderboards/openrouter-rankings/&#34; &gt;OpenRouter rankings&lt;/a&gt; - its sibling usage rankings page, tracked separately.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model provider feature matrix&lt;/a&gt; - gateway features compared across providers.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-selection-for-coding-tasks/&#34; &gt;Model selection for coding tasks&lt;/a&gt; - choosing models once access is settled.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://openrouter.ai/docs/faq&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=openrouter.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://openrouter.ai/docs/faq&lt;/a&gt; - official fee statement (5.5%, $0.80 minimum, crypto 5%), BYOK dollar thresholds, no-markup position, free-tier limits (200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://openrouter.ai/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=openrouter.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://openrouter.ai/pricing&lt;/a&gt; - tier names and fee rows (200, JS-rendered shell; content corroborated via the sources below).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.usagepricing.com/blueprint/openrouter&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.usagepricing.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.usagepricing.com/blueprint/openrouter&lt;/a&gt; - funding history, tier timeline, workspace limits, dated pricing changes (200, facts checked 2026-09-24).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://openrouter.ai/blog/announcements/openrouter-is-joining-stripe/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=openrouter.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://openrouter.ai/blog/announcements/openrouter-is-joining-stripe/&lt;/a&gt; - the joining-Stripe announcement, 2026-08-19: continuity commitments, 10+ trillion tokens per day, 400+ models, 10M+ developers (fetched 200, 2026-09-26).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://ofox.ai/blog/openrouter-pricing-hidden-markup-breakdown-2026/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=ofox.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://ofox.ai/blog/openrouter-pricing-hidden-markup-breakdown-2026/&lt;/a&gt; - independent fee-stack verification, $0.80-minimum math, BYOK metering change (200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.truefoundry.com/blog/openrouter-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.truefoundry.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.truefoundry.com/blog/openrouter-pricing&lt;/a&gt; - critical framing: fee at scale, BYOK, missing SLA (200, published 2026-08-25).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://costgoat.com/pricing/openrouter/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=costgoat.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://costgoat.com/pricing/openrouter/&lt;/a&gt; - per-model price snapshots used for cross-gateway comparison (200, as of 2026-09-25).&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>Qwen Coding Plan</title>
      <link>https://tomrochette.com/agents/model-access/qwen-coding-plan/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/qwen-coding-plan/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>qwen</category><category>subscriptions</category><category>alibaba</category>
      <description>&lt;p&gt;The Qwen Coding Plan is Alibaba Cloud Model Studio&amp;rsquo;s flat-rate subscription that sells Qwen plus rival models (Kimi, GLM, MiniMax) into coding agents through one plan-specific key.&#xA;&lt;strong&gt;It is the only mainstream coding subscription I found that bundles competitors&amp;rsquo; models under one fee, but it carries the strictest automation ban in this category and its product line was restructured twice in six months.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;A subscription from Alibaba Cloud (Model Studio, the international DashScope platform) now sold in two generations.&#xA;The older request-based Coding Plan charges $50/month for Pro and deducts per request; the newer credits-based Token Plan (Singapore region) deducts from a shared Credits pool spanning text, image, video, and speech models.&#xA;Both expose OpenAI-compatible and Anthropic-compatible endpoints and require a plan-specific &lt;code&gt;sk-sp-&lt;/code&gt; key with a dedicated base URL, so one subscription feeds Claude Code, OpenCode, Codex, Cursor, Cline, Qwen Code, Qoder, Kilo CLI, and OpenClaw.&#xA;The Pro roster includes third-party flagships (kimi-k2.5, glm-5, MiniMax-M2.5) alongside qwen3.7-plus, qwen3.6-plus, and the qwen3-coder line.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Actively sold but mid-transition, and the transition is not cosmetic.&#xA;Coding Plan Lite closed to new subscriptions on 2026-03-20 and to renewals and upgrades on 2026-04-13.&#xA;The official Token Plan FAQ states that Coding Plan Pro was a limited-quantity offering that is &amp;ldquo;no longer available once sold out&amp;rdquo;, recommends Token Plan instead, and provides no migration or upgrade path between the two products.&#xA;As of 2026-09-26 the coding-plan docs (updated 2026-09-11) still headline the $50 Pro plan while the Token Plan guide (updated 2026-09-25) sits above it in the docs nav, and Token Plan is Singapore-region only.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;strong&gt;The multi-vendor roster is the real product.&lt;/strong&gt; One key routes to Qwen, Kimi-K2.5, GLM-5, MiniMax-M2.5, and on Token Plan DeepSeek-V4-Pro plus the image, audio, and video families.&lt;/li&gt;&#xA;&lt;li&gt;Token Plan entry is the cheapest on-ramp I verified in this category: $6/month limited-time ($8 list) for 11,500 Credits.&lt;/li&gt;&#xA;&lt;li&gt;$50 Coding Pro bought up to 90,000 requests per month, 6,000 per 5 hours, 45,000 per week, which undercuts pay-as-you-go for anyone spending over $100/month on these APIs.&lt;/li&gt;&#xA;&lt;li&gt;Fixed monthly billing with hard caps: when a window empties, calls pause instead of drawing on your account balance.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;strong&gt;The terms ban automated scripts, CI/CD, batch processing, and application backends, and violations &amp;ldquo;may result in subscription suspension or API key revocation&amp;rdquo;.&lt;/strong&gt;&lt;/li&gt;&#xA;&lt;li&gt;Critical coverage documents enforcement as automated and unpredictable: users report suspensions for rapid-fire long sessions and background-request tools, immediate loss of access for the rest of the billing period, and slow, opaque appeals; as the review puts it, &amp;ldquo;the suspension risk is not theoretical&amp;rdquo;.&lt;/li&gt;&#xA;&lt;li&gt;Quota counts requests, not tokens: Alibaba&amp;rsquo;s own docs say simple tasks burn 5-10 model calls and complex ones 10-30 or more, so 90,000 requests is far fewer agent runs than it sounds.&lt;/li&gt;&#xA;&lt;li&gt;Mixing the plan key with the general pay-as-you-go key produces unexpected API charges, a footgun the official FAQ dedicates a section to.&lt;/li&gt;&#xA;&lt;li&gt;Pro models are pinned to exact snapshots and third-party reviewers report the roster changing with little notice, so a subscription can silently change what it buys.&lt;/li&gt;&#xA;&lt;li&gt;Platform throttling under load is contractually &amp;ldquo;not considered a service interruption or a breach of contract&amp;rdquo;, so slowdowns carry no recourse.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;As of 2026-09-26 two price sheets coexist.&#xA;Coding Plan Pro: $50/month, capped at 6,000 requests per 5 hours, 45,000 per week, and 90,000 per month, whichever hits first (official docs, updated 2026-09-11).&#xA;Token Plan Personal (Singapore region): Lite $6 (list $8), Essential $10 (list $16), Standard $18 (list $25), Pro $68 (list $80) per month for 11,500/25,500/45,000/180,000 Credits, with team seats at $20/$75/$200 and extra bundles at $15 per 20,000 Credits.&#xA;A third-party tracker (data updated 2026-09-24) reports China-side early-bird pricing of ¥39/¥139/¥499 per month, a limited-time night rate of 40% of normal Credits from 22:00, and 88% over-limit billing.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-03&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Lite / Pro&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: Lite listed at ¥40/month in China and promoted as low as $3/month internationally, Pro at $50/month&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;codingplan.org; vibecoding.app&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-03-20&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Lite&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Discontinued for new subscriptions&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Alibaba Cloud coding-plan docs&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-04-13&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Lite&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Renewals and upgrades discontinued&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Alibaba Cloud coding-plan docs&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-08&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Token Plan&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Request-based Coding Plan upgraded to a credits-based Token Plan with China early-bird ¥39/¥139/¥499 per month&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;codingplan.org&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Token Plan (intl)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Lite $6/$8, Essential $10/$16, Standard $18/$25, Pro $68/$80 per month; team seats $20/$75/$200&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Alibaba Cloud Token Plan overview&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;- GLM Coding Plan&lt;/a&gt; wins on quota per dollar; choose Qwen when one subscription covering several vendors&amp;rsquo; models matters more.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/minimax-coding-plan/&#34; &gt;- MiniMax Coding Plan&lt;/a&gt; follows the same China flat-fee pattern with a narrower roster and similar interactivity terms.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/openrouter/&#34; &gt;- OpenRouter&lt;/a&gt; sells the opposite contract: pay per token, no interactivity restrictions, no suspension risk.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for heavy interactive users inside supported harnesses who want Qwen and its rivals under one flat fee and can tolerate mid-cycle throttling.&#xA;Not for CI/CD, batch jobs, unattended agent farms, or anyone who cannot absorb losing access mid-billing-cycle with no guaranteed refund.&#xA;Claim to disagree with: the $50 Coding Plan Pro that Alibaba&amp;rsquo;s docs still headline is now the wrong purchase for almost everyone, because Alibaba&amp;rsquo;s own FAQ calls it a limited offering that is gone once sold out and steers new buyers to the credits-based Token Plan.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created when the owner asked for any remaining subscription providers.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/glm-coding-plan/&#34; &gt;- GLM Coding Plan&lt;/a&gt; - the structural twin: same 5-hour window design, cleaner quota math.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/minimax-coding-plan/&#34; &gt;- MiniMax Coding Plan&lt;/a&gt; - the other China flat-fee plan in this category.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/openrouter/&#34; &gt;- OpenRouter&lt;/a&gt; - the pay-as-you-go alternative without usage-pattern enforcement.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;- Model provider feature matrix&lt;/a&gt; - cross-provider comparison where this plan&amp;rsquo;s rows live.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-selection-for-coding-tasks/&#34; &gt;- Model selection for coding tasks&lt;/a&gt; - how to judge whether Qwen models fit the work.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.alibabacloud.com/help/en/model-studio/coding-plan&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.alibabacloud.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.alibabacloud.com/help/en/model-studio/coding-plan&lt;/a&gt; - official Coding Plan docs: $50 Pro, 6,000/45,000/90,000 request caps, Lite dates of 2026-03-20 and 2026-04-13, automation ban, snapshot-pinned roster, &lt;code&gt;sk-sp-&lt;/code&gt; key rules (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.alibabacloud.com/help/en/model-studio/token-plan-overview&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.alibabacloud.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.alibabacloud.com/help/en/model-studio/token-plan-overview&lt;/a&gt; - official Token Plan overview: USD tiers and Credits, Singapore-only availability, and the FAQ stating Coding Plan Pro was limited-quantity with no migration path (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.alibabacloud.com/en/campaign/ai-scene-coding&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.alibabacloud.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.alibabacloud.com/en/campaign/ai-scene-coding&lt;/a&gt; - official campaign and purchase page: limited-time USD prices, 2X credit promotion, Qwen3.8-Max positioning (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.alibabacloud.com/help/en/model-studio/token-plan-guide&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.alibabacloud.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.alibabacloud.com/help/en/model-studio/token-plan-guide&lt;/a&gt; - official docs hub placing Token Plan above Coding Plan in the product nav as of 2026-09-25 (fetched, HTTP 200).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://codingplan.org/en/plans/qwen&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=codingplan.org&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://codingplan.org/en/plans/qwen&lt;/a&gt; - tracker: credits restructure, ¥39/¥139/¥499 early-bird, 40% night rate from 22:00, legacy ¥40/¥200 China prices (fetched, HTTP 200, data updated 2026-09-24).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://vibecoding.app/blog/alibaba-coding-plan-review&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=vibecoding.app&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://vibecoding.app/blog/alibaba-coding-plan-review&lt;/a&gt; - critical review: suspension reports, opaque appeals, no rollover, roster changes without notice (fetched, HTTP 200, updated 2026-06-17).&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://common-buy-intl.alibabacloud.com/coding-plan&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=common-buy-intl.alibabacloud.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://common-buy-intl.alibabacloud.com/coding-plan&lt;/a&gt; - official purchase URL; fetch reached a 2FA verification wall, so no plan data was read from it (fetch blocked, disclosed).&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>Requesty</title>
      <link>https://tomrochette.com/agents/model-access/requesty/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/requesty/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>ai-gateway</category><category>llm-routing</category><category>eu-data-residency</category>
      <description>&lt;p&gt;Requesty is an EU-hosted managed AI gateway that puts 600+ models behind one OpenAI-compatible endpoint and charges a flat 5% markup on upstream model spend.&lt;/p&gt;&#xA;&lt;p&gt;&lt;strong&gt;The 5% markup is the whole business model: keep your OpenAI SDK code, point the base URL at router.requesty.ai/v1, and pay model cost plus 5% for routing, caching, and governance.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;I read it as pass-through billing with a gateway attached: hosted-only, no subscription tiers, no seat fees, no minimum spend.&#xA;You get routing policies, fallback chains, semantic caching, budget caps, PII detection, and an MCP gateway over one key.&#xA;Their own comparison post says the gateway is built in Rust, and EU residency in Frankfurt is central to their positioning.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Alive and funded: a $3M seed led by 20VC was announced 2025-09-26, with Tapestry VC, Insiders Ventures, and Tiny Supercomputer, and Business Insider covered the pitch deck.&lt;/li&gt;&#xA;&lt;li&gt;The site footer reads © 2026 Requesty Ltd and their blog shipped multiple gateway comparisons through June and September 2026, yet an hn.algolia.com search for &amp;ldquo;requesty.ai&amp;rdquo; returns no story about the company, so their footprint is SEO and docs, not Hacker News.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Pricing is one line long: a $10/1M token model costs $10.50 through them, and the free tier gives 200 requests/day on free models with no credit card.&lt;/li&gt;&#xA;&lt;li&gt;EU (Frankfurt) data residency is included on every plan.&lt;/li&gt;&#xA;&lt;li&gt;Enterprise adds SSO, RBAC, audit logs, guardrails, and service accounts for CI/CD.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;The 5% scales with your spend, so the fee grows exactly when your usage does, unlike flat-priced access.&lt;/li&gt;&#xA;&lt;li&gt;A competitor-authored guide (TrueFoundry, which sells a rival gateway) notes Requesty is hosted-only with no self-hosting, had SOC 2 Type II still in progress (expected Q3 2026), and keeps self-service prompt logging on for up to 30 days by default, with organization-wide zero retention available only on written request.&lt;/li&gt;&#xA;&lt;li&gt;Their own properties disagree on catalog size: 600+ on the pricing page versus 400+ in their June 2026 comparison.&lt;/li&gt;&#xA;&lt;li&gt;The &amp;ldquo;caching makes us net-negative on cost&amp;rdquo; claim (40-60% hit rates) is vendor math I could not verify independently.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Free: $0, free models only, 200 requests/day, routing, caching, spend tracking, and EU residency included.&lt;/li&gt;&#xA;&lt;li&gt;Pay as you go: flat 5% markup on upstream model cost, all 600+ models, budget caps, and MCP gateway, with BYOK carrying 0% markup per their docs (as of 2026-09-26).&lt;/li&gt;&#xA;&lt;li&gt;Enterprise: custom, with SSO (Okta, Azure AD, Google Workspace, custom OIDC), RBAC, approved-model policies, and custom SLAs.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pay as you go&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: flat 5% markup on upstream spend, free tier at 200 requests/day&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;requesty.ai/pricing&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;OpenRouter: a 5.5% fee on credit purchases plus a $0.80 minimum and 365-day credit expiry per Requesty&amp;rsquo;s own comparison; pick OpenRouter for catalog breadth and instant start, Requesty for governance and EU residency.&lt;/li&gt;&#xA;&lt;li&gt;LiteLLM: a free self-hosted proxy with zero markup; pick it if you have DevOps capacity and want no vendor in the billing path.&lt;/li&gt;&#xA;&lt;li&gt;Synthetic (&lt;a href=&#34;https://tomrochette.com/agents/model-access/synthetic/&#34; &gt;Synthetic&lt;/a&gt;): a flat subscription for open-weight coding models, not a general gateway; pick it for agent coding at a fixed monthly cost.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for teams routing production multi-provider traffic who want EU residency, budget governance, and a bill finance can read as a percentage.&#xA;Not for solo developers who only want cheap model access, or anyone who needs self-hosting.&#xA;My most contestable claim: without repetitive traffic that actually hits the cache, you are just paying 5% for a proxy.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/synthetic/&#34; &gt;Synthetic&lt;/a&gt; - the flat-subscription alternative for open-weight coding models.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/opencode/&#34; &gt;OpenCode&lt;/a&gt; - an agent harness commonly pointed at gateways like this one.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model provider feature matrix&lt;/a&gt; - the cross-provider comparison this note feeds.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://requesty.ai/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=requesty.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://requesty.ai/pricing&lt;/a&gt; - 5% markup quote, tier table, 200 requests/day free tier, EU residency (as of 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.requesty.ai/blog/requesty-raises-3m&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.requesty.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.requesty.ai/blog/requesty-raises-3m&lt;/a&gt; - $3M seed led by 20VC, investor list, EU positioning (published 2025-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.requesty.ai/blog/best-llm-routing-platforms-compared-2026-requesty-portkey-litellm-openrouter&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.requesty.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.requesty.ai/blog/best-llm-routing-platforms-compared-2026-requesty-portkey-litellm-openrouter&lt;/a&gt; - their own marketing comparison; Rust, 8ms P50 claim, OpenRouter&amp;rsquo;s 5.5% credit fee (published 2026-06-23)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.truefoundry.com/blog/requesty-ai-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.truefoundry.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.truefoundry.com/blog/requesty-ai-pricing&lt;/a&gt; - competitor-authored critical guide; BYOK 0%, 30-day prompt retention, SOC 2 in progress (published 2026-09-11)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://hn.algolia.com/api/v1/search?query=requesty.ai&amp;amp;tags=story&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=hn.algolia.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://hn.algolia.com/api/v1/search?query=requesty.ai&amp;tags=story&lt;/a&gt; - no meaningful Hacker News footprint found (queried 2026-09-26)&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>SuperGrok</title>
      <link>https://tomrochette.com/agents/model-access/supergrok/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/supergrok/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>grok</category><category>subscriptions</category><category>xai</category>
      <description>&lt;p&gt;SuperGrok is xAI&amp;rsquo;s (SpaceXAI LLC&amp;rsquo;s) consumer subscription ladder for Grok, running from a free tier to $300 per month, which now doubles as the way engineers pay for the Grok Build coding agent.&lt;/p&gt;&#xA;&lt;p&gt;&lt;strong&gt;One subscription, one weekly pool: your Grok chat, your Grok Build runs, and your API calls all drain the same weekly bucket, so the plan you buy for coding is also the plan you spend by chatting.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;A five-step consumer ladder sold by xAI for its Grok products on web, iOS, and Android: Free, SuperGrok Lite, SuperGrok, SuperGrok Plus, and SuperGrok Heavy, with Business and Enterprise seats above.&#xA;The tiers meter chat, Imagine image and video generation, Voice, Grok Bot, and Grok Build, xAI&amp;rsquo;s open-source terminal coding agent, which carries its own row on the official plan comparison for every tier from Free to Enterprise, and the API pricing page separately refers to &amp;ldquo;Grok Build&amp;rsquo;s free tier&amp;rdquo;.&#xA;Above the ladder sits the pay-per-token API, billed separately from any subscription.&#xA;The official pricing page lists dollar figures only for Free, SuperGrok, and Plus; Lite and Heavy show as plan columns without public prices on the web card, so their numbers come from press coverage and third-party guides.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Active and still being restructured: Lite was announced on 2026-03-25 at $10 per month while in testing with selected users, and SuperGrok Plus is absent from a complete tier guide dated 2026-07-06 yet present on the official page on 2026-09-26.&#xA;The defining mechanic is recent and explicit: integration documentation states that every paid Grok subscription gets one weekly usage pool shared across Grok chat, Grok Build, and API access, and that paid usage pauses when the pool empties until the reset time shown in the Usage tab.&#xA;Third-party harnesses can spend the same pool: Warp connects over OAuth and routes Grok models through your xAI account, with its requests labeled API in xAI&amp;rsquo;s usage dashboard.&#xA;&lt;strong&gt;The ladder is converging on one fuzzy weekly pool instead of countable messages, while the tier list itself is still moving underneath it.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;The weekly pool is flexible rather than siloed, so a coding-heavy week can spend the budget on Grok Build and a research-heavy week on DeepSearch and chat.&lt;/li&gt;&#xA;&lt;li&gt;Grok Build is included from Free upward, which makes the $0 tier a working entry point to a frontier coding agent, with paid tiers raising its limits from the same pool.&lt;/li&gt;&#xA;&lt;li&gt;The API escape hatch is cheap for a frontier lab: grok-4.6 and grok-4.7 bill $2.00 input and $6.00 output per million tokens under 200k prompt, with cached input at $0.50, as of 2026-09-26.&lt;/li&gt;&#xA;&lt;li&gt;The subscription travels outside xAI&amp;rsquo;s own apps, since Warp and similar tools authenticate against it directly.&lt;/li&gt;&#xA;&lt;li&gt;Lite at $10 per month would be the cheapest paid on-ramp of any frontier-lab subscription in this category, once it ships generally.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;No quota is published anywhere: no messages, no tokens, no pool size, so the only way to size a plan is to burn a month and read the Usage tab.&lt;/li&gt;&#xA;&lt;li&gt;The shared pool cuts both ways: an afternoon of video generation or idle chat quietly deletes tomorrow&amp;rsquo;s coding budget.&lt;/li&gt;&#xA;&lt;li&gt;Lite is not confirmed as generally available, announced as a test with selected users in March 2026, and Heavy&amp;rsquo;s $300 figure rests on third-party guides because the official web card does not expose it.&lt;/li&gt;&#xA;&lt;li&gt;Coding is the ladder&amp;rsquo;s weakest argument: the critical guide I fetched concludes Claude and ChatGPT lead on coding help and recommends SuperGrok for live X data and image work, not for code.&lt;/li&gt;&#xA;&lt;li&gt;Imagine, the media product the higher tiers are priced around, is the surface under active lawsuits, a US congressional inquiry, and EU, UK, and Canada probes, per the same critical guide.&lt;/li&gt;&#xA;&lt;li&gt;A subscription connected through a third-party harness carries no zero-data-retention guarantee from that harness, since retention on xAI&amp;rsquo;s side is governed by your own xAI account.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Free: $0, limited usage, Grok Build included with limits.&#xA;SuperGrok Lite: $10/month, announced 2026-03-25, basic creation tools, 480p video up to 6 seconds, 2x longer chats than Free, one AI agent, in testing at announcement.&#xA;SuperGrok: $30/month, Grok 4.6, higher rate limits across all features, image and video generation.&#xA;SuperGrok Plus: $100/month, everything in SuperGrok plus 1080p video, significantly higher usage across Chat, Imagine, Voice, and Build, priority access at peak times.&#xA;SuperGrok Heavy: $300/month per third-party guides as of 2026-07-06, the multi-agent Grok 4 Heavy tier.&#xA;API escape hatch, billed separately: grok-4.6 and grok-4.7 at $2.00/$6.00 per 1M tokens under 200k prompt ($4.00/$12.00 above), $0.50 cached input; grok-build-0.1 at $1.00/$2.00; Grok 4.7 Fast at 2x rates, exclusive to Cursor and Grok Build and excluded from Grok Build&amp;rsquo;s free tier.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-01-24&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;SuperGrok / Heavy&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Baseline: SuperGrok $30/month and Heavy $300/month, no Lite or Plus tier; the July 2026 update describes these two anchor prices as unchanged&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://aitoolanalysis.com/supergrok-subscription-price-2026/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=aitoolanalysis.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://aitoolanalysis.com/supergrok-subscription-price-2026/&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-03-25&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;SuperGrok Lite&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Announced at $10/month by Musk on X, in testing with selected users, global rollout promised later in the year&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://www.businesstoday.in/technology/story/xai-makes-grok-affordable-with-new-supergrok-lite-plan-check-price-and-what-it-offers-522462-2026-03-26&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.businesstoday.in&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.businesstoday.in/technology/story/xai-makes-grok-affordable-with-new-supergrok-lite-plan-check-price-and-what-it-offers-522462-2026-03-26&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;SuperGrok Plus&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Observed at $100/month on the official pricing page, absent from the 2026-07-06 complete tier guide, so added in the July to September window&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;&lt;a href=&#34;https://x.ai/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=x.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://x.ai/pricing&lt;/a&gt;&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/claude-plans/&#34; &gt;Claude plans&lt;/a&gt;: the $20 to $200 Anthropic ladder that runs Claude Code; it matches SuperGrok at the $100 rung and carries the stronger coding reputation, while publishing no more about its quotas than xAI does.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chatgpt-plans/&#34; &gt;ChatGPT plans&lt;/a&gt;: the rival consumer ladder at $20 to $200, which publishes per-model message ranges, a discipline xAI has not adopted.&lt;/li&gt;&#xA;&lt;li&gt;The &lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model Provider Feature Matrix&lt;/a&gt;: the per-token path, and the reason the shared pool matters less if your coding runs on the API anyway.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for engineers already inside the Grok and X ecosystem who want one bill across chat and Grok Build and can tolerate an unpublished weekly pool.&#xA;Not for coding-first engineers choosing a primary agent subscription, where Claude plans or the GLM Coding Plan buy more verified coding throughput per dollar.&#xA;My disagreeable claim: SuperGrok Plus at $100 is a media-creator tier wearing an engineer&amp;rsquo;s price tag, because on a shared pool the extra $70 mostly buys 1080p video and peak-hour priority, not more code.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created when the owner asked for any remaining subscription providers.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/claude-plans/&#34; &gt;Claude plans&lt;/a&gt; - the incumbent subscription this ladder prices against, rung for rung.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/chatgpt-plans/&#34; &gt;ChatGPT plans&lt;/a&gt; - the other big-lab consumer ladder, facing the same unpublished-quota problem.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/grok-build/&#34; &gt;Grok Build&lt;/a&gt; - the harness these subscriptions meter, including the wire-level privacy analysis of its default path.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-provider-feature-matrix/&#34; &gt;Model Provider Feature Matrix&lt;/a&gt; - xAI&amp;rsquo;s API-side row, where the $2/$6 escape hatch lives.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://x.ai/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=x.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://x.ai/pricing&lt;/a&gt; - official plan cards (Free $0, SuperGrok $30, Plus $100) and the plan comparison carrying Lite, Heavy, and Grok Build rows (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://docs.warp.dev/agents/inference/grok-subscription&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=docs.warp.dev&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://docs.warp.dev/agents/inference/grok-subscription&lt;/a&gt; - the one weekly pool shared across Grok chat, Grok Build, and API access, pause-on-empty behavior, third-party-harness ZDR caveat (fetched 200, page updated 2026-09-24)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://docs.x.ai/developers/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=docs.x.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://docs.x.ai/developers/pricing&lt;/a&gt; - API escape hatch: grok-4.6/4.7 at $2/$6 under 200k prompt, $0.50 cached, $4/$12 above, grok-build-0.1 at $1/$2, Grok 4.7 Fast exclusivity (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.businesstoday.in/technology/story/xai-makes-grok-affordable-with-new-supergrok-lite-plan-check-price-and-what-it-offers-522462-2026-03-26&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.businesstoday.in&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.businesstoday.in/technology/story/xai-makes-grok-affordable-with-new-supergrok-lite-plan-check-price-and-what-it-offers-522462-2026-03-26&lt;/a&gt; - Lite announcement: $10/month, testing phase with selected users, 480p 6-second video, 2x chats, one agent (fetched 200, 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://aitoolanalysis.com/supergrok-subscription-price-2026/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=aitoolanalysis.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://aitoolanalysis.com/supergrok-subscription-price-2026/&lt;/a&gt; - critical guide: full tier table including Heavy $300, the coding-benchmarks caveat, and the safety litigation section (fetched 200, updated 2026-07-06)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://codeagentswarm.com/en/guides/grok-build-pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=codeagentswarm.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://codeagentswarm.com/en/guides/grok-build-pricing&lt;/a&gt; - independent Grok Build plan-table guide, attempted four times this run and rate-limited (429) on every attempt, so nothing here is cited from it&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
    <item>
      <title>Synthetic</title>
      <link>https://tomrochette.com/agents/model-access/synthetic/</link>
      <pubDate>Sat, 26 Sep 2026 00:00:00 +0000</pubDate>
      <author>tom@tomrochette.com (Tom Rochette)</author>
      <guid>https://tomrochette.com/agents/model-access/synthetic/</guid>
      <category>research-note</category><category>agent-curated</category><category>fully-ai-generated</category><category>llm=glm-5.3-flash</category><category>model-access</category><category>open-weight-models</category><category>llm-subscription</category><category>coding-agents</category>
      <description>&lt;p&gt;Synthetic (synthetic.new) sells flat monthly subscriptions and pay-per-token billing for open-weight LLMs served on its own US and EU infrastructure, aimed squarely at coding-agent users.&lt;/p&gt;&#xA;&lt;p&gt;&lt;strong&gt;It is the closest thing to Claude Pro for open-weight models: $30/month for 500 requests per 5 hours across models like DeepSeek-V4.1-Flash, GLM-5.3-Flash, and Kimi-K3, with prompts never stored longer than 14 days.&lt;/strong&gt;&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;What it is&#xA;    &lt;div id=&#34;what-it-is&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#what-it-is&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;A privacy-focused inference provider that runs open-weight models itself rather than reselling closed APIs, with &amp;ldquo;always-on&amp;rdquo; models included in subscriptions and usage-based billing for enterprise and on-demand needs.&#xA;Any OpenAI-compatible tool works against api.synthetic.new/v1 (Anthropic-compatible /messages endpoints exist too), and OpenCode ships Synthetic as a native provider you select with /connect.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Status&#xA;    &lt;div id=&#34;status&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#status&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Launched via Show HN on 2025-08-28 (31 points, 21 comments), with the founder answering questions in-thread.&lt;/li&gt;&#xA;&lt;li&gt;By September 2025 third parties described it as 19 always-on models for $20-60/month; as of 2026-09-26 the site markets a single $30/month pack plus usage-based billing.&lt;/li&gt;&#xA;&lt;li&gt;Distribution is real, with OpenCode, Kilo Code, Crush, Claude Code, GitHub Copilot, and OpenClaw all listed in its docs, and its marketing has also gotten visible enough to attract pushback, including one 2026 Reddit thread warning users away and another accusing its community of astroturfing.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Strengths&#xA;    &lt;div id=&#34;strengths&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#strengths&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Flat cost with forgiving metering: requests are counted per API call, so a parallel tool-call batch is one request, per the founder on HN.&lt;/li&gt;&#xA;&lt;li&gt;Privacy is specific rather than vibes: API prompts and completions cannot be stored longer than 14 days and only for debugging, and Kilo&amp;rsquo;s docs confirm no training on your data.&lt;/li&gt;&#xA;&lt;li&gt;syn: aliases (syn:large:text) auto-route to the latest recommended model, protecting you from pinning names that later 404.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Cautions&#xA;    &lt;div id=&#34;cautions&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#cautions&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;A 2026 r/opencodeCLI thread titled &amp;ldquo;Stay away from synthetic.new&amp;rdquo; reports the 3x-Claude-limits framing did not hold in practice, with limits hitting much sooner than expected, possibly from inefficient tool calling in the Chinese open models.&lt;/li&gt;&#xA;&lt;li&gt;One pack allows only 1 concurrent request per model; you must buy more packs to raise parallelism, which matters for multi-agent setups.&lt;/li&gt;&#xA;&lt;li&gt;Model rotation is a documented risk: their own docs warn that pinned model names will eventually 404.&lt;/li&gt;&#xA;&lt;li&gt;An r/kimi thread from mid-2026 accuses Synthetic community moderators of coordinated affiliate-link promotion; I cannot verify the accusation, so I treat surrounding hype with care.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Pricing&#xA;    &lt;div id=&#34;pricing&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#pricing&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Subscription pack: $30/month ($1/day) for 500 requests per 5 hours, advertised as 3x the rate limits of Claude&amp;rsquo;s $20/month plan, with 1 concurrent request per model and UI plus API access (as of 2026-09-26).&lt;/li&gt;&#xA;&lt;li&gt;Usage-based: pay-per-token on always-on models and pay-per-minute on-demand, pitched at enterprise.&lt;/li&gt;&#xA;&lt;li&gt;All always-on models plus embeddings are included in every subscription, and embeddings do not count against the rate limit.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Price history&#xA;    &lt;div id=&#34;price-history&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#price-history&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;table&gt;&#xA;&#x9;&lt;thead&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Date&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Plan&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Change&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;th&gt;Source&lt;/th&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/thead&gt;&#xA;&#x9;&lt;tbody&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2025-10-08&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Standard&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;$20/month, 135 messages per 5 hours&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Wayback capture of synthetic.new/pricing&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2025-10-08&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pro&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;$60/month, 1,350 messages per 5 hours&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Wayback capture of synthetic.new/pricing&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&#x9;&#x9;&lt;tr&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;2026-09-26&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;Pack (x1)&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;$30/month ($1/day), 500 requests per 5 hours, 1 concurrent request per model; the exact date the $20/$60 tiers ended is not in my sources&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&#x9;&#x9;&lt;td&gt;synthetic.new/pricing&lt;/td&gt;&#xA;&#x9;&#x9;&#x9;&lt;/tr&gt;&#xA;&#x9;&lt;/tbody&gt;&#xA;&lt;/table&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Compared to&#xA;    &lt;div id=&#34;compared-to&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#compared-to&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;Requesty (&lt;a href=&#34;https://tomrochette.com/agents/model-access/requesty/&#34; &gt;../requesty/index.md&lt;/a&gt;): a general gateway metered at 5% of spend; pick Synthetic when you want a fixed bill and open-weight models only.&lt;/li&gt;&#xA;&lt;li&gt;Claude Pro / Claude Code: 1.5x the price ($30 vs $20) for closed frontier models; pick Claude for peak quality, Synthetic for privacy and flat cost.&lt;/li&gt;&#xA;&lt;li&gt;Direct open-model hosts (DeepSeek, Z.ai): pay per token and often cheaper at low volume; Synthetic wins when agent request volume would blow past token budgets.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Bottom line&#xA;    &lt;div id=&#34;bottom-line&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#bottom-line&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;p&gt;Recommended for OpenCode or Kilo users who hit rate limits daily, want a predictable $30/month, and prefer prompts not retained beyond 14 days.&#xA;Not for anyone needing frontier closed models, guaranteed capacity, or high concurrency on a single pack.&#xA;My most contestable claim: the &amp;ldquo;3x Claude limits&amp;rdquo; line is the weakest reason to buy, because the most detailed public user report says real-world limits feel closer to Claude&amp;rsquo;s than the marketing implies.&lt;/p&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;Changes&#xA;    &lt;div id=&#34;changes&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#changes&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;2026-09-26 - Created.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;See also&#xA;    &lt;div id=&#34;see-also&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#see-also&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/model-access/requesty/&#34; &gt;Requesty&lt;/a&gt; - the percentage-markup gateway alternative in this same category.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/opencode/&#34; &gt;OpenCode&lt;/a&gt; - the harness that ships Synthetic as a native provider.&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://tomrochette.com/agents/harnesses/kilo-code/&#34; &gt;Kilo Code&lt;/a&gt; - documents Synthetic as a first-class provider.&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;&#xA;&lt;h2 class=&#34;relative group&#34;&gt;References&#xA;    &lt;div id=&#34;references&#34; class=&#34;anchor&#34;&gt;&lt;/div&gt;&#xA;    &#xA;    &lt;span&#xA;        class=&#34;absolute top-0 w-6 transition-opacity opacity-0 -start-6 not-prose group-hover:opacity-100 select-none&#34;&gt;&#xA;        &lt;a class=&#34;text-primary-300 dark:text-neutral-700 !no-underline&#34; href=&#34;#references&#34; aria-label=&#34;Anchor&#34;&gt;#&lt;/a&gt;&#xA;    &lt;/span&gt;&#xA;    &#xA;&lt;/h2&gt;&#xA;&lt;ul&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://synthetic.new/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=synthetic.new&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://synthetic.new/pricing&lt;/a&gt; - current $30/month pack, 500 requests/5hr, 1 concurrent request per model, usage-based option (as of 2026-09-26; JS-rendered page, text extracted from fetched HTML)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://web.archive.org/web/20251008095830id_/https://synthetic.new/pricing&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=web.archive.org&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://web.archive.org/web/20251008095830id_/https://synthetic.new/pricing&lt;/a&gt; - October 2025 pricing: Standard $20/135 messages per 5h, Pro $60/1,350&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://hn.algolia.com/api/v1/items/45055763&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=hn.algolia.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://hn.algolia.com/api/v1/items/45055763&lt;/a&gt; - Show HN launch thread; founder on request counting, 14-day retention, and GLM-4.5 as flagship (2025-08-28)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://dev.synthetic.new/docs/api/models&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=dev.synthetic.new&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://dev.synthetic.new/docs/api/models&lt;/a&gt; - current always-on model list, syn: aliases, model-rotation warning (as of 2026-09-26)&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://dev.synthetic.new/docs/guides/opencode&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=dev.synthetic.new&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://dev.synthetic.new/docs/guides/opencode&lt;/a&gt; - native OpenCode provider setup&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://kilo.ai/docs/ai-providers/synthetic&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=kilo.ai&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://kilo.ai/docs/ai-providers/synthetic&lt;/a&gt; - Kilo integration, no-training and 14-day auto-delete privacy claims&lt;/li&gt;&#xA;&lt;li&gt;&lt;a href=&#34;https://www.reddit.com/r/opencodeCLI/comments/1rfdadw/stay_away_from_syntheticnew/&#34;  target=&#34;_blank&#34; rel=&#34;noreferrer&#34;&gt;&lt;img class=&#34;external-link-favicon&#34; src=&#34;https://www.google.com/s2/favicons?domain=www.reddit.com&amp;sz=128&#34; alt=&#34;&#34; width=&#34;16&#34; height=&#34;16&#34; loading=&#34;lazy&#34;&gt;https://www.reddit.com/r/opencodeCLI/comments/1rfdadw/stay_away_from_syntheticnew/&lt;/a&gt; - critical user report on real-world limits; direct fetch returned 403, content recovered via search-result extraction&lt;/li&gt;&#xA;&lt;/ul&gt;&#xA;</description>
      
    </item>
    
  </channel>
</rss>
