Qwen3.8 Max Prime
Last verified October 3, 2026 · Alibaba pricing ↗$4.00/1M input · $12.00/1M output · 1.0M context · Alibaba
Count Tokens
—
Tokens
—
Input Cost
—
Output Cost
Estimate Monthly Cost
Monthly Cost Estimator
Quick:
<$0.0001/mo
Pick a preset above or enter custom usage
Alternatives to Qwen3.8 Max Prime
Pricing Details
AlibabaAlibaba Model Studio 'Prime mode' (announced Yunqi conference 2026-09-22, on OpenRouter 2026-09-23): same Qwen3.8 Max weights on faster serving, Alibaba claims 1.5-2x the standard API's TPS. OpenRouter card $4.00 input / $0.50 cache read / $12.00 output, exactly 2x standard Qwen3.8 Max ($2.00 / $0.25 / $6.00); no cache write listed. Alibaba's own price list shows it in China (Beijing) only at CNY 24 / CNY 72 ($3.301 / $9.902), also 2x the Beijing standard rate. Measured on OpenRouter (read 2026-10-03): p50 70 tok/s and 962 ms to first token at 06:06 UTC (329 requests), 58 tok/s and 885 ms at 06:13 UTC (618 requests), vs standard 36 tok/s and 3,914 ms, so 1.6-1.9x throughput for 2x price. Context 1M, max output 131,072, text/image/video input. No Artificial Analysis page yet; standard Qwen3.8 Max scores 45.
Input / 1M tokens
$4
Output / 1M tokens
$12
Context Window
1.0M
Max Output
131K
Price History
Launched at current price on 2026-09-22. No price changes recorded since.
Frequently Asked Questions
Common questions about Qwen3.8 Max Prime pricing and usage