Model Spend Arena2ND ED.
serves endpoints · 2026-08-14

Fireworks

Official site · 🇺🇸 United States

Models it serves

10 endpoints, cheapest first. This is the view the model-by-model tables cannot give you: what one provider actually offers, and at what precision.

ModelQuant$/M in$/M outContextUptime 30m
gpt-oss-20bunknown$0.07$0.30131,07298.9%
DeepSeek V4 Flashunknown$0.14$0.281,048,57696.4%
MiniMax M2.7unknown$0.30$1.20196,608
Muse Glimmerunknown$0.35$1.50131,072100.0%
Kimi K2.6unknown$0.95$4.00262,144
DeepSeek V4 Prounknown$1.32$3.961,048,576100.0%
GLM-5.2unknown$1.40$4.401,048,57698.7%
GLM-5.2unknown$2.10$6.601,048,57698.5%
Kimi K3unknown$3.00$15.001,048,57699.6%
Kimi K3unknown$4.50$22.501,048,57699.6%
◆ Some links on this page are referral links: if you sign up through them this site may earn a commission, at no extra cost to you. Rankings, prices and measurements are taken from the sources listed under Method, and are not influenced by whether a provider has a referral programme.