Models /Qwen3 VL 8B Thinking /Where to run
Provider guide

Where to run Qwen3 VL 8B Thinking

3 live offers tracked — output prices vary 1.5× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.12 / $1.36
per 1M tokens in / out
FASTEST MEASURED
138 tok/s
measured throughput · $0.18 input / $2.10 output per million tokens
Alibaba CloudCHEAPEST$0.12 in/1M$1.36 out/1M132 tok/s131K7d agoOpenRouter$0.18 in/1M$2.10 out/1M131K4h agoAlibaba Cloud$0.18 in/1M$2.10 out/1M138 tok/s131Kfp84h ago
Full specs, hardware verdicts and benchmarks on theQwen3 VL 8B Thinking model page →