Provider guide
Where to run Qwen3 VL 8B Instruct
4 live listings tracked — output prices vary 1.6× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.
FASTEST MEASURED
55 tok/s
measured throughput · $0.12 input / $0.46 output per million tokens
OpenRouterCHEAPEST$0.12 in/1M$0.46 out/1M—262K—2 hours agoAlibaba Cloud$0.12 in/1M$0.46 out/1M55 tok/s131K—2 hours agoNovita AI$0.080 in/1M$0.50 out/1M—131K—21 days agoParasailzero-retention$0.25 in/1M$0.75 out/1M13 tok/s262Kbf162 hours ago
Full specs, hardware verdicts and benchmarks on the Qwen3 VL 8B Instruct model page →