Provider guide
Where to run Qwen3.5 397B A17B
15 live offers tracked — output prices vary 1.5× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.
FASTEST MEASURED
63 tok/s
measured throughput · $0.39 input / $2.34 output per million tokens
OpenRouterCHEAPEST$0.39 in/1M$2.34 out/1M—262K—5h agoAlibaba Cloud$0.39 in/1M$2.34 out/1M63 tok/s262K—6d agoAlibaba Cloud$0.39 in/1M$2.34 out/1M34 tok/s262Kfp85h agoDeepInfra$0.45 in/1M$3.00 out/1M—262Kfp85h agoChutes$0.45 in/1M$3.00 out/1M5 tok/s262Kfp85h agoAtlasCloud$0.55 in/1M$3.50 out/1M35 tok/s262Kfp85h agoNovita AI$0.60 in/1M$3.60 out/1M—262K—5h agoGMICloud$0.60 in/1M$3.60 out/1M—262Kfp87h agoStreamLake$0.60 in/1M$3.60 out/1M9 tok/s256K—5h agoDigitalOcean Gradientzero-retention$0.39 in/1M$2.45 out/1M4 tok/s131K—5h agoDeepInfrazero-retention$0.45 in/1M$3.00 out/1M27 tok/s262Kfp87h agoPhalazero-retention$0.55 in/1M$3.50 out/1M48 tok/s262K—5h agoNovita AIzero-retention$0.60 in/1M$3.60 out/1M47 tok/s262K—5h agoParasailzero-retention$0.50 in/1M$3.60 out/1M49 tok/s262Kfp85h agoVenice AIzero-retention$0.75 in/1M$4.50 out/1M49 tok/s128K—5h ago
Full specs, hardware verdicts and benchmarks on theQwen3.5 397B A17B model page →