Models /Llama 3.1 70B Instruct /Where to run
Provider guide

Where to run Llama 3.1 70B Instruct

5 live offers tracked.

CHEAPEST
$0.40 / $0.40
per 1M tokens in / out
FASTEST MEASURED
27 tok/s
measured throughput · $0.80 input / $0.80 output per million tokens
OpenRouterCHEAPEST$0.40 in/1M$0.40 out/1M131K3h agoDeepInfra$0.40 in/1M$0.40 out/1M23 tok/s131Kfp86d agoDeepInfraturbo tier$0.40 in/1M$0.40 out/1M19 tok/s131Kfp83h agoAmazon Bedrockzero-retention$0.72 in/1M$0.72 out/1M21 tok/s131K3h agoCoreWeavezero-retention$0.80 in/1M$0.80 out/1M27 tok/s128Kbf163h ago
Full specs, hardware verdicts and benchmarks on theLlama 3.1 70B Instruct model page →