Skip to main content
TokenCost logoTokenCost

TokenBlog

Model releases, pricing breakdowns, and practical guides for developers.

Model ReleaseSeptember 2, 2026
Claude
Latest·12 min read

Claude Fable 5.1 charges what Fable 5 charged on every line of the rate card except one…

Anthropic shipped Fable 5.1 and Mythos 5.1 on September 1 and left almost the whole card alone: $10.00 input, $50.00 output, $12.50 and $20.00 cache writes, $5.00 and $25.00 on Batch, a…

ComparisonAugust 30, 2026
Grok

Four companies sell Grok 4.6 and all four print $2.00 and $6.00. The whole difference…

xAI, Google, Amazon and Microsoft all list Grok 4.6 at $2.00 per million input tokens, $0.50 cached and $6.00 output. Copy those three…

11 min read
Model ReleaseAugust 29, 2026
Model Release

Tencent priced its new flagship at 6 yuan per million input tokens and printed $0.834 next…

Hy4 preview landed on Thursday at 06:09 UTC: 770B parameters with 49B active, a 1,048,576-token window, Apache 2.0 weights, and one seller…

12 min read
GuideAugust 28, 2026
Guide

Eight AI API discounts expire on a published date between September 7 and January 1. Half…

Introductory pricing has become the normal way to launch a model, so a growing share of the rates in your spreadsheet are offers with a…

14 min read
ComparisonAugust 27, 2026
GLM

Two models called Flash arrived five and a half hours apart yesterday. One of them looks…

Z.ai listed GLM-5.3-Flash at 13:59 UTC on August 26 and Alibaba listed Qwen3.8-Flash at 19:37. Every tracker shows the first at $0.075 per…

12 min read
ComparisonAugust 26, 2026
Comparison

Three models charge exactly $2.00 per million input tokens. The same 4K screenshot costs…

Every pricing page quotes dollars per million tokens and almost none of them say how many tokens a picture is. We collected the counting…

12 min read
ComparisonAugust 25, 2026
Comparison

Localizing a million characters into four languages costs $0.48 on Tencent's new…

Tencent put its Hy-MT2 translation family on a metered API last week, three months after releasing the weights under Apache 2.0. We counted…

11 min read
Model ReleaseAugust 24, 2026
Model Release

Ox Alpha bills $0 today. Its listing carries a setting that only two of OpenRouter's 422…

An unattributed model appeared on OpenRouter on August 20 at 20:04 UTC priced at nothing, and the weekend went to guessing which lab owns…

11 min read
IndustryAugust 23, 2026
GPT

OpenAI cut GPT-5.6 Sol to exactly 0.8 times Claude Opus 5 on every line it sells. Azure…

OpenAI dropped GPT-5.6 Sol on August 21 under a headline of "over 20%", which undersells it on output and oversells it as permanent. Input…

11 min read
Model ReleaseAugust 22, 2026
GLM

GLM-5.3 bills the same $1.40 and $4.40 that GLM-5.2 and GLM-5.1 billed, and finishing a…

Z.ai put a per-token price on GLM-5.3 on August 18, four days after the model arrived reachable only through a subscription, and the price…

11 min read
GuideAugust 21, 2026
Claude

Claude's new browser tool bills about 6,610 input tokens before it reads a word of your…

Anthropic made computer use generally available on August 19 and launched a browser tool beside it, publishing the token overhead for both…

11 min read
GuideAugust 18, 2026
GPT

Bedrock sells GPT-5.6 Sol at exactly OpenAI's list price and also at 1.10x it. Which one…

AWS publishes five Bedrock model cards for GPT-5.6 and only three of them have a cheap row. Ordinary GPT-5.6 Sol carries three routing…

10 min read
ResearchAugust 17, 2026
Research

Terminal-Bench 3.0 publishes what every run cost, which almost no other leaderboard does…

Agent leaderboards almost never tell you what the run cost. This one has a COST column and a TOKENS column next to the score, split per row…

11 min read
IndustryAugust 16, 2026
DeepSeek

DeepSeek said the increase would be significant. It is 1.80x off-peak and 3.59x at peak…

At 16:00 UTC today DeepSeek's rate card split in two. V4-Flash goes from a flat $0.14 and $0.28 per million tokens to $0.22 and $0.66 for…

10 min read
Model ReleaseAugust 15, 2026
Grok

Grok 4.6 costs exactly what Grok 4.5 costs on input, on output, and at the cliff. The only…

xAI shipped Grok 4.6 on August 12 and put it on the rate card of the model it succeeds. Input $2.00, output $6.00, a 500,000 token window…

11 min read
Model ReleaseAugust 14, 2026
Gemini

Google cut Gemini 3.6 Flash in half on Thursday and launched Gemini 3.7 Flash at the same…

Gemini 3.7 Flash went GA on August 13 at $0.75 input and $3.75 output per million tokens, and the coverage all repeated the same framing…

12 min read
Model ReleaseAugust 11, 2026
Model Release

Meta gave Muse Glimmer's weights away yesterday and exactly one company sells it. $0.35…

Meta shipped a 30B dense model under Apache 2.0 on August 10, its first open weights in more than a year, and named twelve launch partners…

11 min read
GuideAugust 10, 2026
Guide

Imagen 4 stops answering on August 17. Google's pages send you to two different…

Three endpoints go dark next Monday, and the retirement changes the billing unit rather than just the number. Imagen 4 sold an image for a…

13 min read
IndustryAugust 9, 2026
GPT

On August 1 we wrote that nobody was going to bill you $20 per million input tokens. Four…

OpenAI's changelog for August 5 is one sentence long and mentions no price: Fast mode now supports long-context requests for GPT-5.6 Sol…

11 min read
GuideAugust 8, 2026
Claude

Anthropic will now stop an agent session at a number you choose, written in whole cents as…

Session budgets shipped for Claude Managed Agents on August 7. Set budget.max_list_cost.amount to "2500" and the session stops at $25, which…

12 min read
IndustryAugust 8, 2026
DeepSeek

DeepSeek says a significant price increase is coming and will not say how significant. We…

A three-sentence notice went up on DeepSeek's pricing page on August 6, as footnote number two under the rate table. No percentage, no…

11 min read
Model ReleaseAugust 7, 2026
Model Release

Ling-3.0-flash sells for $0.021, $0.06 and $0.075 per million input tokens right now. Same…

Ant Group put a 124B MoE on the API on July 23 and did not open the weights until August 2, one day before the free trial expired. What…

12 min read
IndustryAugust 6, 2026
Industry

Meta will cut your API bill 22x for permission to train on your prompts. The rate that…

Muse Spark 1.2 shipped on August 5 with a second model ID beside it. muse-spark-1.2-contributor bills $0.10 input, $0.002 cached input and…

11 min read
Model ReleaseAugust 5, 2026
Grok

grok-voice-latest moved to Think Fast 2.0 this morning, so unchanged code now bills 60%…

xAI repointed the grok-voice-latest alias today, taking anyone who used it from $0.05 a minute of audio to $0.08. That is $3.00 an hour…

13 min read
Model ReleaseAugust 4, 2026
Qwen

Qwen3.8-Max prices output 60% under Kimi K3. On the same benchmark suite the bill came in…

Alibaba took its 2.4T-parameter flagship to general availability on August 3 and every gateway now quotes $2 input and $6 output per…

11 min read