TokenBlog
Model releases, pricing breakdowns, and practical guides for developers.
Claude Fable 5.1 charges what Fable 5 charged on every line of the rate card except one…
Anthropic shipped Fable 5.1 and Mythos 5.1 on September 1 and left almost the whole card alone: $10.00 input, $50.00 output, $12.50 and $20.00 cache writes, $5.00 and $25.00 on Batch, a…
Four companies sell Grok 4.6 and all four print $2.00 and $6.00. The whole difference…
xAI, Google, Amazon and Microsoft all list Grok 4.6 at $2.00 per million input tokens, $0.50 cached and $6.00 output. Copy those three…
11 min readTencent priced its new flagship at 6 yuan per million input tokens and printed $0.834 next…
Hy4 preview landed on Thursday at 06:09 UTC: 770B parameters with 49B active, a 1,048,576-token window, Apache 2.0 weights, and one seller…
12 min readEight AI API discounts expire on a published date between September 7 and January 1. Half…
Introductory pricing has become the normal way to launch a model, so a growing share of the rates in your spreadsheet are offers with a…
14 min readTwo models called Flash arrived five and a half hours apart yesterday. One of them looks…
Z.ai listed GLM-5.3-Flash at 13:59 UTC on August 26 and Alibaba listed Qwen3.8-Flash at 19:37. Every tracker shows the first at $0.075 per…
12 min readThree models charge exactly $2.00 per million input tokens. The same 4K screenshot costs…
Every pricing page quotes dollars per million tokens and almost none of them say how many tokens a picture is. We collected the counting…
12 min readLocalizing a million characters into four languages costs $0.48 on Tencent's new…
Tencent put its Hy-MT2 translation family on a metered API last week, three months after releasing the weights under Apache 2.0. We counted…
11 min readOx Alpha bills $0 today. Its listing carries a setting that only two of OpenRouter's 422…
An unattributed model appeared on OpenRouter on August 20 at 20:04 UTC priced at nothing, and the weekend went to guessing which lab owns…
11 min readOpenAI cut GPT-5.6 Sol to exactly 0.8 times Claude Opus 5 on every line it sells. Azure…
OpenAI dropped GPT-5.6 Sol on August 21 under a headline of "over 20%", which undersells it on output and oversells it as permanent. Input…
11 min readGLM-5.3 bills the same $1.40 and $4.40 that GLM-5.2 and GLM-5.1 billed, and finishing a…
Z.ai put a per-token price on GLM-5.3 on August 18, four days after the model arrived reachable only through a subscription, and the price…
11 min readClaude's new browser tool bills about 6,610 input tokens before it reads a word of your…
Anthropic made computer use generally available on August 19 and launched a browser tool beside it, publishing the token overhead for both…
11 min readBedrock sells GPT-5.6 Sol at exactly OpenAI's list price and also at 1.10x it. Which one…
AWS publishes five Bedrock model cards for GPT-5.6 and only three of them have a cheap row. Ordinary GPT-5.6 Sol carries three routing…
10 min readTerminal-Bench 3.0 publishes what every run cost, which almost no other leaderboard does…
Agent leaderboards almost never tell you what the run cost. This one has a COST column and a TOKENS column next to the score, split per row…
11 min readDeepSeek said the increase would be significant. It is 1.80x off-peak and 3.59x at peak…
At 16:00 UTC today DeepSeek's rate card split in two. V4-Flash goes from a flat $0.14 and $0.28 per million tokens to $0.22 and $0.66 for…
10 min readGrok 4.6 costs exactly what Grok 4.5 costs on input, on output, and at the cliff. The only…
xAI shipped Grok 4.6 on August 12 and put it on the rate card of the model it succeeds. Input $2.00, output $6.00, a 500,000 token window…
11 min readGoogle cut Gemini 3.6 Flash in half on Thursday and launched Gemini 3.7 Flash at the same…
Gemini 3.7 Flash went GA on August 13 at $0.75 input and $3.75 output per million tokens, and the coverage all repeated the same framing…
12 min readMeta gave Muse Glimmer's weights away yesterday and exactly one company sells it. $0.35…
Meta shipped a 30B dense model under Apache 2.0 on August 10, its first open weights in more than a year, and named twelve launch partners…
11 min readImagen 4 stops answering on August 17. Google's pages send you to two different…
Three endpoints go dark next Monday, and the retirement changes the billing unit rather than just the number. Imagen 4 sold an image for a…
13 min readOn August 1 we wrote that nobody was going to bill you $20 per million input tokens. Four…
OpenAI's changelog for August 5 is one sentence long and mentions no price: Fast mode now supports long-context requests for GPT-5.6 Sol…
11 min readAnthropic will now stop an agent session at a number you choose, written in whole cents as…
Session budgets shipped for Claude Managed Agents on August 7. Set budget.max_list_cost.amount to "2500" and the session stops at $25, which…
12 min readDeepSeek says a significant price increase is coming and will not say how significant. We…
A three-sentence notice went up on DeepSeek's pricing page on August 6, as footnote number two under the rate table. No percentage, no…
11 min readLing-3.0-flash sells for $0.021, $0.06 and $0.075 per million input tokens right now. Same…
Ant Group put a 124B MoE on the API on July 23 and did not open the weights until August 2, one day before the free trial expired. What…
12 min readMeta will cut your API bill 22x for permission to train on your prompts. The rate that…
Muse Spark 1.2 shipped on August 5 with a second model ID beside it. muse-spark-1.2-contributor bills $0.10 input, $0.002 cached input and…
11 min readgrok-voice-latest moved to Think Fast 2.0 this morning, so unchanged code now bills 60%…
xAI repointed the grok-voice-latest alias today, taking anyone who used it from $0.05 a minute of audio to $0.08. That is $3.00 an hour…
13 min readQwen3.8-Max prices output 60% under Kimi K3. On the same benchmark suite the bill came in…
Alibaba took its 2.4T-parameter flagship to general availability on August 3 and every gateway now quotes $2 input and $6 output per…
11 min read




































































































































































