Provider guide
Where to run GLM 5
10 live listings tracked — output prices vary 1.7× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.
FASTEST MEASURED
77 tok/s
measured throughput · $1.00 input / $3.20 output per million tokens
OpenRouterCHEAPEST$0.60 in/1M$1.92 out/1M—205K—3 hours agoStreamLake$0.60 in/1M$1.92 out/1M41 tok/s198Kfp83 hours agoGMICloud$0.60 in/1M$1.92 out/1M66 tok/s203Kfp83 hours agoDeepInfra$0.60 in/1M$2.08 out/1M—203Kfp428 days agoBaidu$0.70 in/1M$2.24 out/1M48 tok/s203Kfp83 hours agoSiliconFlowzero-retention$0.95 in/1M$2.55 out/1M34 tok/s205Kfp83 hours agoVenice AIzero-retention$1.00 in/1M$3.20 out/1M58 tok/s198Kfp83 hours agoNovita AIzero-retention through OpenRouterDirect and through OpenRouter$1.00 in/1M$3.20 out/1M48 tok/sthrough OpenRouter203Kfp89 hours ago direct3 hours ago through OpenRouterAmazon Bedrockzero-retention$1.00 in/1M$3.20 out/1M77 tok/s203K—3 hours agoZ.AIzero-retention$1.00 in/1M$3.20 out/1M59 tok/s203Kfp83 hours ago
Full specs, hardware verdicts and benchmarks on the GLM 5 model page →