Model Spend Arena2ND ED.
serves endpoints Β· 2026-10-03

SiliconFlow

Official site Β· πŸ‡¨πŸ‡³ China

Models it serves

25 endpoints, cheapest first. This is the view the model-by-model tables cannot give you: what one provider actually offers, and at what precision.

ModelQuant$/M in$/M outContextUptime 30m
gpt-oss-20bfp8$0.04$0.18131,07297.0%
Qwen3.5-9Bfp8$0.10$0.15262,14499.9%
Gemma 4 26B A4B fp8$0.14$0.40262,144100.0%
GLM-5.3-Flashfp8$0.15$0.501,048,57697.6%
Qwen3 32Bfp8$0.14$0.57131,072100.0%
gpt-oss-120bfp8$0.15$0.60131,07276.7%
DeepSeek V3.2fp8$0.26$0.42163,84099.6%
DeepSeek V4 Flashfp8$0.22$0.661,048,57699.4%
DeepSeek V3 0324fp8$0.25$1.00163,84099.9%
DeepSeek V3.1 Terminusfp8$0.27$1.00163,84099.4%
DeepSeek V4.1 Flashfp8$0.30$1.201,048,57699.9%
Qwen3.6 35B A3Bfp8$0.24$1.80262,144100.0%
Qwen3.5-35B-A3Bfp8$0.24$1.80262,144β€”
DeepSeek V4 Flash Visionfp8$0.44$1.321,048,576100.0%
Qwen3.5-122B-A10Bfp8$0.26$2.08262,14496.1%
Gemma 4 31Bfp8$0.75$1.00262,14439.4%
Kimi K2.5int4$0.45$2.25262,144100.0%
Qwen3.6 27Bfp8$0.30$3.20262,144β€”
GLM-5.2fp8$0.70$2.201,048,576100.0%
GLM-5.3fp8$0.70$2.201,048,576100.0%
Kimi K2.6fp8$0.77$3.40262,144100.0%
Kimi K2.7 Codefp8$0.86$3.80262,144β€”
GLM-5.1fp8$1.19$3.74204,800100.0%
DeepSeek V4 Profp8$1.32$3.961,048,576β€”
Qwen3.8 2.4T A95Bfp8$2.00$6.001,048,576100.0%