Provider guide
Where to run Qwen3 Next 80B A3B Thinking
6 live offers tracked — output prices vary 1.9× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.
FASTEST MEASURED
228 tok/s
measured throughput · $0.098 input / $0.78 output per million tokens
Alibaba Cloud$0.098 in/1M$0.78 out/1M228 tok/s131K—7d agoOpenRouterCHEAPEST$0.15 in/1M$1.20 out/1M—262K—44m agoAlibaba Cloud$0.15 in/1M$1.20 out/1M218 tok/s131Kfp842m agoNovita AI$0.15 in/1M$1.50 out/1M—131K—7h agoGoogle Vertex AIzero-retention$0.15 in/1M$1.20 out/1M38 tok/s262K—42m agoNebius AI Studiozero-retention$0.15 in/1M$1.20 out/1M107 tok/s128Kfp842m ago
Full specs, hardware verdicts and benchmarks on theQwen3 Next 80B A3B Thinking model page →