origin not confirmed
3 endpoints, cheapest first. This is the view the model-by-model tables cannot give you: what one provider actually offers, and at what precision.
| Model | Quant | $/M in | $/M out | Context | Uptime 30m |
|---|---|---|---|---|---|
| gpt-oss-120b | fp8 | $0.08 | $0.50 | 131,072 | 99.3% |
| DeepSeek V4 Flash | fp8 | $0.17 | $0.50 | 1,048,576 | 97.6% |
| GLM 4.7 | fp4 | $0.60 | $2.50 | 131,072 | 98.4% |