Model Spend Arena2ND ED.
serves endpoints · 2026-10-03

CoreWeave

Official site · origin not confirmed

Models it serves

16 endpoints, cheapest first. This is the view the model-by-model tables cannot give you: what one provider actually offers, and at what precision.

ModelQuant$/M in$/M outContextUptime 30m
gpt-oss-20bfp4$0.03$0.13131,072100.0%
gpt-oss-120bfp4$0.03$0.17131,07298.8%
Nemotron 3.5 Lightningbf16$0.07$0.20262,144100.0%
Granite 4.2 8Bbf16$0.10$0.15131,072100.0%
Gemma 4 26B A4B bf16$0.10$0.30262,144100.0%
Gemma 4 31Bfp4$0.10$0.34262,14498.3%
DeepSeek V4 Flashfp8$0.13$0.28262,144100.0%
GLM-5.3-Flashnvfp4$0.15$0.501,048,57699.9%
DeepSeek V4.1 Flashfp8$0.20$0.651,048,57699.2%
MiniMax-M3fp4$0.23$0.96262,14499.8%
Qwen3.6 35B A3Bfp8$0.25$1.25262,144100.0%
Qwen3.8-27Bfp8$0.40$3.00262,144100.0%
GLM-5.2fp4$0.76$2.421,048,576—
Kimi K2.6fp4$0.65$3.41262,14499.9%
Kimi K2.7 Codeint4$0.71$3.50262,144100.0%
DeepSeek V4 Profp8$1.31$3.961,048,576100.0%