Models /Qwen3 235B A22B Instruct 2507 /Where to run
Provider guide

Where to run Qwen3 235B A22B Instruct 2507

10 live listings tracked — output prices vary 5.2× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.22 / $0.88
per 1M tokens in / out
FASTEST MEASURED
43 tok/s
measured throughput · $0.15 input / $0.60 output per million tokens
GMICloud$0.087 in/1M$0.35 out/1M34 tok/s262Kfp82 hours agoDeepInfrazero-retention through OpenRouterDirect and through OpenRouter$0.090 in/1M$0.55 out/1M16 tok/sthrough OpenRouter262Kfp88 hours ago direct2 hours ago through OpenRouterNovita AIzero-retention through OpenRouterDirect and through OpenRouter$0.090 in/1M$0.58 out/1M31 tok/sthrough OpenRouter131Kfp88 hours ago direct2 hours ago through OpenRouterAlibaba Cloud$0.15 in/1M$0.60 out/1M43 tok/s131K—2 hours agoNebius AI Studiozero-retention$0.20 in/1M$0.60 out/1M23 tok/s262Kfp82 hours agoVenice AIzero-retention$0.15 in/1M$0.75 out/1M12 tok/s128Kfp82 hours agoParasailzero-retention$0.14 in/1M$0.80 out/1M24 tok/s131Kfp82 hours agoStreamLake$0.21 in/1M$0.84 out/1M32 tok/s128K—2 hours agoGoogle Vertex AIzero-retentionCHEAPEST$0.22 in/1M$0.88 out/1M16 tok/s262K—2 hours agoOpenRouter$0.46 in/1M$1.82 out/1M—131K—5 days ago
Full specs, hardware verdicts and benchmarks on the Qwen3 235B A22B Instruct 2507 model page →