Cheapest GLM 4.6 API providers

13 providers host GLM 4.6. The cheapest right now is Venice at $0.43/MTok input and $1.75/MTok output. Prices below are pulled live from each provider, updated daily.

What would GLM 4.6 cost you?

SiliconFlow is 45% cheaper than Io Net at this workload.
Input tokens / month150M
Output tokens / month30M

Projected monthly cost = (input price × 150M) + (output price × 30M). Drag the sliders to match your actual workload; the chart re-ranks live.

#HostContextInput $/MTokOutput $/MTokBlendedUptime 30mQuant
1Venice198k$0.43$1.75$0.6599.87%fp4
2Venice198k$0.43$1.75$0.65100.00%fp4
3SiliconFlow205k$0.39$1.90$0.6499.43%fp8
4DeepInfra203k$0.50$2.00$0.75100.00%fp4
5DeepInfra203k$0.50$2.00$0.7599.62%fp4
6Novita205k$0.55$2.20$0.8395.05%bf16
7Novita205k$0.55$2.20$0.8396.20%bf16
8BaseTen200k$0.60$2.20$0.87fp4
9AtlasCloud203k$0.60$2.20$0.87fp8
10Z Ai203k$0.60$2.20$0.8797.16%fp4
11AtlasCloud203k$0.60$2.20$0.87fp8
12Z Ai203k$0.60$2.20$0.8781.04%fp4
13Io Net131k$0.85$2.75$1.17fp8

Other rankings for GLM 4.6