Models /Qwen3 32B /Where to run
Provider guide

Where to run Qwen3 32B

8 live offers tracked — output prices vary 1.6× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.080 / $0.28
per 1M tokens in / out
FASTEST MEASURED
361 tok/s
measured throughput · $0.29 input / $0.59 output per million tokens
OpenRouterCHEAPEST$0.080 in/1M$0.28 out/1M131K59m agoDeepInfra$0.080 in/1M$0.28 out/1M41Kfp87h agoNebius AI Studio$0.10 in/1M$0.30 out/1M23 tok/s41Kfp87h agoNebius AI Studio$0.10 in/1M$0.30 out/1M25 tok/s41Kfp86d agoNovita AI$0.10 in/1M$0.45 out/1M41K7h agoDeepInfrazero-retention$0.080 in/1M$0.28 out/1M31 tok/s41Kfp858m agoSiliconFlowzero-retention$0.14 in/1M$0.57 out/1M16 tok/s131Kfp858m agoGroqzero-retention$0.29 in/1M$0.59 out/1M361 tok/s131K58m ago
Full specs, hardware verdicts and benchmarks on theQwen3 32B model page →