llama-3.1-8b-instruct API price — 4 gateways compared
Cheapest right now: Novita at $0.02/MTok in — 2.5× cheaper than the priciest gateway.
from $0.02 per
million input tokens · 2.5× spread between gateways
· updated continually
Gateway $ in /MTok $ out /MTok
Context Input price history Model ID
Novita $0.02 $0.05 16,384 meta-llama/llama-3.1-8b-instruct use OpenRouter fees $0.05 $0.08 131,072 meta-llama/llama-3.1-8b-instruct use Requesty $0.05 $0.05 16,384 novita/meta-llama/llama-3.1-8b-instruct use Glama $0.05 $0.08 128,000 meta/llama-3.1-8b-instruct use
Fees beyond the token price OpenRouter : No markup on token prices (provider pass-through), but credit purchases cost 5.5% via card ($0.80 minimum) or 5% via crypto; BYOK usage above the monthly allowance is charged 5%. source
Changelog for llama-3.1-8b-instruct
2026-08-27 requesty raised llama-3.1-8b-instruct-output 11.1%: 0.045 → 0.05 usd_per_mtok
2026-08-27 requesty raised llama-3.1-8b-instruct-input 11.1%: 0.045 → 0.05 usd_per_mtok
llama-3.1-8b-instruct pricing FAQ
How much does llama-3.1-8b-instruct cost per million tokens? The cheapest tracked gateway charges $0.02 per million input tokens and $0.05 per million output tokens (Novita). Prices refresh continually.
Which gateway is cheapest for llama-3.1-8b-instruct? Novita for input tokens and Novita for output tokens, among the 4 gateways we track.
Do llama-3.1-8b-instruct prices differ between gateways? Yes — the most expensive tracked gateway charges 2.5× the cheapest for input tokens. Identical model, identical tokens, different bill.
← All LLM prices · API
Every datum links its source and is versioned in
public git history .
Data: CC BY 4.0 — cite gpudiff.com. Nameplate specs are vendor claims, not measured
throughput. Vast prices are the 25th percentile of verified marketplace listings
(what a careful buyer actually gets); RunPod prices are list. Outbound provider
links may carry referral codes and sponsored units are labelled; neither ever
affects the numbers. Sponsor this site .