Provider guide
Where to run GLM 4.7
8 live listings tracked.
FASTEST MEASURED
35 tok/s
measured throughput · $0.60 input / $2.20 output per million tokens
DeepInfrazero-retention through OpenRouterDirect and through OpenRouter$0.40 in/1M$1.75 out/1M33 tok/sthrough OpenRouter203Kfp45 hours ago direct17 hours ago through OpenRouterAtlasCloud$0.52 in/1M$1.85 out/1M—203Kfp87 days agoVenice AIzero-retention$0.40 in/1M$1.93 out/1M26 tok/s198Kfp417 hours agoNovita AIzero-retention through OpenRouterDirect and through OpenRouter$0.60 in/1M$0.54 in/1M$2.20 out/1Mdirect$1.98 out/1Mthrough OpenRouter26 tok/sthrough OpenRouter205Kfp85 hours agoZ.AIzero-retention$0.60 in/1M$2.20 out/1M26 tok/s203Kfp45 hours agoGoogle Vertex AIzero-retention$0.60 in/1M$2.20 out/1M35 tok/s200K—5 hours agoOpenRouterCHEAPEST$0.60 in/1M$2.20 out/1M—205K—5 hours agoMancer 2zero-retention$0.70 in/1M$2.50 out/1M21 tok/s131Kfp417 hours ago
Full specs, hardware verdicts and benchmarks on the GLM 4.7 model page →