Models /Llama 3.2 3B Instruct /Where to run
Provider guide

Where to run Llama 3.2 3B Instruct

4 live offers tracked — output prices vary 6.7× between the cheapest and most expensive host, so the provider choice matters as much as the model choice.

CHEAPEST
$0.050 / $0.33
per 1M tokens in / out
FASTEST MEASURED
73 tok/s
measured throughput · $0.051 input / $0.34 output per million tokens
Novita AI$0.030 in/1M$0.050 out/1M33K9h agoOpenRouterCHEAPEST$0.050 in/1M$0.33 out/1M131K3h agoCloudflare Workers AI$0.051 in/1M$0.34 out/1M73 tok/s80K3h agoParasailzero-retention$0.050 in/1M$0.33 out/1M61 tok/s131Kbf169h ago
Full specs, hardware verdicts and benchmarks on theLlama 3.2 3B Instruct model page →