Provider guide
Where to run Qwen3 VL 235B A22B Thinking
5 live offers tracked — output prices vary 1.5× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.
FASTEST MEASURED
60 tok/s
measured throughput · $0.40 input / $4.00 output per million tokens
Alibaba CloudCHEAPEST$0.26 in/1M$2.60 out/1M—131K—7d agoNovita AI$0.98 in/1M$3.95 out/1M—131K—7h agoOpenRouter$0.40 in/1M$4.00 out/1M—131K—32h agoAlibaba Cloud$0.40 in/1M$4.00 out/1M60 tok/s131Kfp843m agoNovita AIzero-retention$0.98 in/1M$3.95 out/1M41 tok/s131Kbf1643m ago
Full specs, hardware verdicts and benchmarks on theQwen3 VL 235B A22B Thinking model page →