Models /Llama 4 Scout /Where to run
Provider guide

Where to run Llama 4 Scout

4 live listings tracked — output prices vary 2.3× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.10 / $0.30
per 1M tokens in / out
FASTEST MEASURED
102 tok/s
measured throughput · $0.25 input / $0.70 output per million tokens
OpenRouterCHEAPEST$0.10 in/1M$0.30 out/1M—1.3M—30 min agoDeepInfrazero-retention through OpenRouterDirect and through OpenRouter$0.10 in/1M$0.30 out/1M29 tok/sthrough OpenRouter328Kfp86 hours ago direct29 min ago through OpenRouterNovita AIzero-retention through OpenRouterDirect and through OpenRouter$0.18 in/1M$0.59 out/1M16 tok/sthrough OpenRouter131Kbf166 hours ago direct29 min ago through OpenRouterGoogle Vertex AIzero-retention$0.25 in/1M$0.70 out/1M102 tok/s1.3M—29 min ago
Full specs, hardware verdicts and benchmarks on the Llama 4 Scout model page →