Model Spend Arena2ND ED.
builds models · serves endpoints · 2026-10-03

Alibaba / Qwen

Official site · 🇨🇳 China

Alibaba / Qwen sells 2 different plans, easy to confuse. The Coding Plan is for agentic coding only and is metered in requests; the Token Plan is metered in credits and covers every modality. Pick the Coding Plan if you live in a coding agent and want a flat monthly cap; pick the Token Plan (Individual for one person, Team for seats) if you also do image/video or want the cheaper $6-$68 entry points. Neither converts cleanly to tokens. And mind which link you sign up through: the promotional price (50% new-user coupon, cheaper Team seats) only applies from specific campaign pages, so the same plan can cost different amounts by the door you enter. One non-price difference matters most: Alibaba lists "no model training of your content" as a Team-only feature, so on the cheaper Individual tiers your interactions can be used to train models.

Jump to: Coding Plan (Pro) (requests · coding only) · Token Plan (Individual + Team) (credits · all modalities)

Coding Plan (Pro)

Coding Plan (Pro)

Flat monthly cap for agentic coding, metered in requests — not tokens. One live tier: Pro at $50/month.

Buy — Model Studio →

Tier$/monthIncluded usage
Pro$50.00—

console only partly convertible

As advertised. 6,000 requests / 5h · 45,000 / week · 90,000 / month

What that actually means. Request caps are published (6,000/5h) but no per-request cost, so requests do not convert to tokens. Feature request to show remaining was closed as not planned. El 50% de descuento de nuevo usuario NO se confirma (2026-09-11): el doc oficial no menciona ninguna promocion y hay indicios de que el cupon expiro el 2026-04-01. La puerta firme es el doc, a $50 pelados.

Where to check what is left. home.qwencloud.com/billing/coding-plan

Models covered. qwen3-coder-plus, qwen3-coder-next, qwen3.7-plus, qwen3.6-plus, qwen3.5-plus, qwen3-max-2026-01-23, kimi-k2.5, glm-5, glm-4.7, MiniMax-M2.5

A SEPARATE product from the Token Plan below, and easy to confuse with it. The Coding Plan is agentic-coding only (Claude Code / Qwen Code / Cline style), metered in REQUESTS, and now has a single live tier: Pro at $50/month. The Lite tier was discontinued for new subs 2026-03-20 (renewals ended 2026-04-13), though comparison sites still list it. Docs say a coding task uses ~5-10 requests (simple) to 30+ (complex), so requests are not tasks and do not convert to tokens. 2026-10-02 (doc del 2026-09-28): PLAZAS LIMITADAS, dato nuevo: 'Slots are limited and available on a first-come, first-served basis. New slots are restocked daily at 00:00:00 (UTC+08:00)'. Publica tambien como se repone la cuota: la de 5h rueda minuto a minuto, la semanal resetea los lunes a las 00:00 UTC+8 y la mensual en la fecha de renovacion.

checked 2026-10-02 · high · source

Token Plan (Individual + Team)

Token Plan (Individual + Team)

Credits across every modality (text/image/video/audio). Buy it as an Individual (Lite/Standard/Pro, $6–$68/mo) or per Team seat. Credits do not convert to tokens — the exchange rate is unpublished.

Token Plan — Individual (Lite $6, 50% new-user coupon) → · Token Plan — Team (asientos $20/$75/$200) →

Tier$/monthIncluded quotaConcurrent
Individual Lite$6.0011,5001-2
Individual Essential$10.0025,5002-3
Individual Standard$18.0045,0003-4
Individual Pro$68.00180,0006-8
Team Standard seat$20-3025,000/seat—
Team Pro seat$75-100100,000/seat—
Team Max seat$200.00250,000/seat—

console only units unstated

As advertised. Individual credits (subida publicada el 2026-10-02): Lite 11.500/mes · Essential 25.500 (2,2x Lite) · Standard 45.000 (3,9x Lite) · Pro 180.000 (15,7x Lite). Team per seat: Standard 25.000/mes · Pro 100.000 · Max 250.000. Agentes simultaneos (Individual): 1-2 / 2-3 / 3-4 / 6-8. Las ventanas de 5h y 7 dias NO las publica ninguna pagina oficial: lo que habia era de terceros y se retiro.

What that actually means. Metered in CREDITS with a deliberately unpublished exchange rate. Neither Alibaba page states how many tokens a credit buys, a per-model credit-consumption rate, or a dollar value for a credit -- an independent decode concluded the plan is "priced in a currency whose exchange rate is unpublished", so it cannot be turned into a token quota. The credits are metered on rolling 5h and 7-day windows as well as monthly (so the monthly figure is a ceiling you cannot spend in a burst). A limited-time promo gives Qwen3.8-Max and DeepSeek-V4-Pro-0813 2x usage per credit (era 10x y solo Qwen3.8-Max hasta 2026-09), which only underlines that a hidden per-credit usage rate exists. Distinct from the request-metered Coding Plan above: different product, different unit, all modalities (text/image/video/audio) not just coding. A 50% new-user coupon and a separate $2-off coupon apply only from specific links, and a limited-time promo can 10x usage per credit -- so the same plan can cost materially different amounts depending on the URL you sign up through. 2026-08-28: RESOLVED the Team-seat cross-page ambiguity by rendering ai-token-plan directly with a real browser (loaded cleanly, no captcha this time -- likely session/cookie dependent, since the same URL returned a slider challenge to a fresh headless session earlier the same day). This single page shows Standard Seat $20/mo (struck-through original $30), Pro Seat $75/mo (original $100), Max Seat $200/mo flat -- i.e. the DISCOUNTED figures, not $30/$100/$200 as previously attributed to this page. All individual-tier credit multiples also confirmed exactly against the live page: Standard = 4x Lite, Pro = 16x Lite; Team Pro = 4x Standard, Team Max = 10x Standard. Did not attempt ai-landing-page-token again since ai-token-plan alone now settles the number.

Where to check what is left. alibabacloud.com model studio console

Models covered. qwen3.8-max-preview, qwen3.7-max, qwen3.7-plus, qwen3.6-flash, glm-5.2, deepseek-v4-pro, kimi, MiniMax-M2.5, (+ Wan / HappyHorse / Qwen-image: image/video/audio)

2026-10-02: TRAMO NUEVO, Individual Essential a $10 al mes ($100 al ano) con 25.500 creditos, 2,2 veces el Lite. Y los tres tramos viejos SUBEN de creditos: Lite de ~10.000 a 11.500, Standard de ~40.000 a 45.000 y Pro de ~160.000 a 180.000. Los precios tachados de hoy son $8/$16/$25/$80 en Individual y $30/$100 en Team. Su propia tabla comparativa sigue sin enterarse del Essential.
Individual vs Team — read before the price. Individual and Team are not just a price difference. Alibaba's own Individual-vs-Team comparison lists FOUR things as Team-only: enterprise-grade data security, enterprise management, unified agent/usage support, and — the one that matters most — "no model training of your content". Listed as a Team benefit, it means the Individual (Personal) plans carry no such guarantee: on the cheap $6-$68 tiers, your interactions can be used to train models. Team also adds a shared quota pack across seats. So the real trade is privacy and pooled quota (Team, from $20/seat) versus lowest entry price (Individual, from $6) — not just capacity.

checked 2026-10-02 · high (tiers/credits as listed) / the credit unit is undefined by the vendor · source 1 source 2 source 3

Models it builds

ModelIntelligencetok/s$/M in$/M out$ / sessionCache
Qwen3.7 Max29.5—$1.48$4.42$0.920288.8%
Qwen3.7 Plus25.255.7$0.32$1.28$0.200788.3%
Qwen3.6 Plus27.0—$0.33$1.95——
Qwen3.6 27B21.4—$0.32$3.20——
Qwen3.5 397B A17B21.485.5$0.55$3.50——
Qwen3.5-122B-A10B17.7143.3$0.26$2.08——
Qwen3.6 35B A3B18.2134.9$0.15$1.00——
Qwen3.5-35B-A3B19.3—$0.15$1.00——
Qwen3 Coder Next9.283.7$0.12$0.80——
Qwen3.5-9B13.381.4$0.10$0.15——
Qwen3.8 Max45.437.4$2.00$6.00——
Qwen3.8-27B33.745.8$0.42$3.00——
Qwen3 32B8.6—$0.08$0.28——
Qwen3 14B8.2—$0.12$0.24——
Qwen3 8B7.3—$0.12$0.45——
Qwen3 235B A22B 250712.7—$0.45$1.82——
Qwen3 30B A3B 25079.8—$0.12$0.50——
Qwen3.8 2.4T A95B39.938.6$2.00$6.00——

Models it serves

26 endpoints, cheapest first. This is the view the model-by-model tables cannot give you: what one provider actually offers, and at what precision.

ModelQuant$/M in$/M outContextUptime 30m
Qwen3 8Bunknown$0.12$0.45131,072100.0%
Qwen3 30B A3B 2507unknown$0.13$0.52131,072100.0%
Qwen3 14Bunknown$0.23$0.91131,072100.0%
Qwen3.5-35B-A3Bunknown$0.16$1.30262,14499.9%
DeepSeek V4.1 Flashunknown$0.30$1.201,000,000100.0%
DeepSeek V4 Flashunknown$0.35$1.061,000,00099.1%
DeepSeek V3.2fp8$0.37$1.11131,07296.9%
Qwen3.7 Plusunknown$0.32$1.281,000,000100.0%
Qwen3 Coder Nextunknown$0.30$1.50262,144—
Qwen3.5-122B-A10Bunknown$0.26$2.08262,14499.2%
Qwen3.6 Plusunknown$0.33$1.951,000,000100.0%
Qwen3 235B A22B 2507unknown$0.45$1.82131,07297.1%
Qwen3.5 397B A17Bunknown$0.39$2.34262,14499.1%
Qwen3.8-27Bunknown$0.42$2.551,000,00099.8%
Qwen3.6 27Bunknown$0.45$2.70262,14499.3%
GLM-5.2fp8$0.97$3.041,000,000100.0%
DeepSeek V4 Prounknown$1.12$3.371,000,00095.4%
Kimi K2.7 Codefp8$0.95$4.00262,144—
GLM-5.3unknown$1.19$3.741,000,00099.8%
GLM-5.1fp8$1.33$4.18202,74599.3%
Qwen3.7 Maxunknown$1.48$4.421,000,000100.0%
Qwen3.8 Maxunknown$2.00$6.001,000,00099.8%
Qwen3.8 2.4T A95Bunknown$2.00$6.001,000,000100.0%
GLM-5.2fp8$2.31$7.261,000,000—
GLM-5.3unknown$2.80$8.801,000,000—
Kimi K3unknown$3.45$17.251,048,57698.7%

Where to buy its models

Everyone serving these models, cheapest first, with the numeric precision each one runs. The lab is often not the cheapest place to buy its own model.

ModelServed byQuant$/M in$/M outUptime 30m
Qwen3.7 MaxAlibaba / Qwenunknown$1.48$4.42100.0%
Qwen3.7 PlusAlibaba / Qwenunknown$0.32$1.28100.0%
Qwen3.6 PlusAlibaba / Qwenunknown$0.33$1.95100.0%
Qwen3.6 27BChutesfp8$0.30$2.00100.0%
Qwen3.6 27BPhalaunknown$0.32$2.7099.3%
Qwen3.6 27BAlibaba / Qwenunknown$0.45$2.7099.3%
Qwen3.6 27BSiliconFlowfp8$0.30$3.20—
Qwen3.6 27BDeepInfrafp8$0.32$3.20100.0%
Qwen3.6 27BVenicefp8$0.33$3.25—
Qwen3.5 397B A17BAlibaba / Qwenunknown$0.39$2.3499.1%
Qwen3.5 397B A17BDeepInfrafp8$0.45$3.00100.0%
Qwen3.5 397B A17BParasailfp8$0.50$3.6099.9%
Qwen3.5 397B A17BAtlasCloudfp8$0.55$3.5095.8%
Qwen3.5 397B A17BDigitalOceanunknown$0.55$3.5096.9%
Qwen3.5 397B A17BPhalaunknown$0.55$3.5098.7%
Qwen3.5 397B A17BGMICloudfp8$0.60$3.6099.4%
Qwen3.5 397B A17BNovitaunknown$0.60$3.6089.6%
Qwen3.5 397B A17BStreamLakeunknown$0.60$3.6097.5%
Qwen3.5 397B A17BVeniceunknown$0.75$4.5098.5%
Qwen3.5-122B-A10BAlibaba / Qwenunknown$0.26$2.0899.2%
Qwen3.5-122B-A10BSiliconFlowfp8$0.26$2.0896.1%
Qwen3.5-122B-A10BAtlasCloudfp8$0.30$2.4099.3%
Qwen3.5-122B-A10BNovitabf16$0.40$3.2099.1%
Qwen3.6 35B A3BDarkbloomfp4$0.05$0.70100.0%
Qwen3.6 35B A3BAkashMLfp8$0.10$0.90100.0%
Qwen3.6 35B A3BDeepInfrafp8$0.10$0.9599.9%
Qwen3.6 35B A3BVenicefp8$0.10$1.00100.0%
Qwen3.6 35B A3BParasailfp8$0.15$1.00100.0%
Qwen3.6 35B A3BAtlasCloudfp8$0.19$1.1198.3%
Qwen3.6 35B A3BPhalaunknown$0.20$1.27100.0%
Qwen3.6 35B A3BCoreWeavefp8$0.25$1.25100.0%
Qwen3.6 35B A3BSiliconFlowfp8$0.24$1.80100.0%
Qwen3.5-35B-A3BDarkbloomfp4$0.08$0.75100.0%
Qwen3.5-35B-A3BDeepInfrafp8$0.14$1.00100.0%
Qwen3.5-35B-A3BParasailfp8$0.15$1.00100.0%
Qwen3.5-35B-A3BAlibaba / Qwenunknown$0.16$1.3099.9%
Qwen3.5-35B-A3BVeniceunknown$0.31$1.25100.0%
Qwen3.5-35B-A3BAtlasCloudfp8$0.22$1.80—
Qwen3.5-35B-A3BSiliconFlowfp8$0.24$1.80—
Qwen3 Coder NextParasailbf16$0.12$0.80100.0%
Qwen3 Coder NextStreamLakeunknown$0.18$0.9098.9%
Qwen3 Coder NextNovitafp8$0.20$1.50—
Qwen3 Coder NextAlibaba / Qwenunknown$0.30$1.50—
Qwen3.5-9BDarkbloomfp4$0.08$0.13100.0%
Qwen3.5-9BDeepInfrabf16$0.10$0.15100.0%
Qwen3.5-9BSiliconFlowfp8$0.10$0.1599.9%
Qwen3.5-9BVenicefp8$0.10$0.15100.0%
Qwen3.5-9BParasailbf16$0.10$0.25100.0%
Qwen3.5-9BTogetherunknown$0.17$0.2599.9%
Qwen3.8 MaxAlibaba / Qwenunknown$2.00$6.0099.8%
Qwen3.8-27BDeepInfrabf16$0.15$1.8899.7%
Qwen3.8-27BPhalaunknown$0.15$1.8898.8%
Qwen3.8-27BDarkbloomfp4$0.05$2.20100.0%
Qwen3.8-27BAkashMLfp8$0.20$1.78100.0%
Qwen3.8-27BIonstreamfp8$0.09$2.2099.6%
Qwen3.8-27BChutesfp8$0.24$2.2099.2%
Qwen3.8-27BParasailfp8$0.24$2.20100.0%
Qwen3.8-27BMancer 2fp8$0.20$2.5099.8%
Qwen3.8-27BDekaLLMunknown$0.05$3.0099.8%
Qwen3.8-27BAlibaba / Qwenunknown$0.42$2.5599.8%
Qwen3.8-27BCoreWeavefp8$0.40$3.00100.0%
Qwen3.8-27BNovitaunknown$0.42$3.0098.8%
Qwen3.8-27BWaferunknown$0.02$4.35100.0%
Qwen3.8-27BCerebrasfp16$0.99$1.49100.0%
Qwen3.8-27BCloudflareunknown$0.45$3.2087.7%
Qwen3.8-27BVenicefp8$0.45$3.2099.4%
Qwen3 32BDeepInfrafp8$0.08$0.28100.0%
Qwen3 32BSiliconFlowfp8$0.14$0.57100.0%
Qwen3 14BNextBitint4$0.10$0.22100.0%
Qwen3 14BDeepInfrafp8$0.12$0.2497.0%
Qwen3 14BAlibaba / Qwenunknown$0.23$0.91100.0%
Qwen3 8BAlibaba / Qwenunknown$0.12$0.45100.0%
Qwen3 235B A22B 2507Alibaba / Qwenunknown$0.45$1.8297.1%
Qwen3 30B A3B 2507DeepInfrafp8$0.12$0.50100.0%
Qwen3 30B A3B 2507Alibaba / Qwenunknown$0.13$0.52100.0%
Qwen3.8 2.4T A95BAlibaba / Qwenunknown$2.00$6.00100.0%
Qwen3.8 2.4T A95BDeepInfrafp4$2.00$6.00—
Qwen3.8 2.4T A95BModalunknown$2.00$6.00—
Qwen3.8 2.4T A95BNovitaunknown$2.00$6.00100.0%
Qwen3.8 2.4T A95BSiliconFlowfp8$2.00$6.00100.0%
Qwen3.8 2.4T A95BTogetherunknown$2.00$6.00100.0%
Qwen3.8 2.4T A95BVeniceunknown$2.00$6.00—