Provider guide
Where to run GLM 4.6
5 live listings tracked.
FASTEST MEASURED
11 tok/s
measured throughput · $0.50 input / $2.00 output per million tokens
Venice AIzero-retention$0.43 in/1M$1.75 out/1M5 tok/s198Kfp428 min agoOpenRouterCHEAPEST$0.43 in/1M$1.75 out/1M—205K—30 min agoDeepInfrazero-retention through OpenRouterDirect and through OpenRouter$0.50 in/1M$2.00 out/1M11 tok/sthrough OpenRouter203Kfp46 hours ago direct28 min ago through OpenRouterNovita AIzero-retention through OpenRouterDirect and through OpenRouter$0.55 in/1M$2.20 out/1M1 tok/sthrough OpenRouter205Kbf166 hours ago direct28 min ago through OpenRouterZ.AIzero-retention$0.60 in/1M$2.20 out/1M1 tok/s203Kfp428 min ago
Full specs, hardware verdicts and benchmarks on the GLM 4.6 model page →