gpudiff the public record of change in GPU cloud pricing
Follow the diffs: RSS · Free API · Sponsor this site

llama-3.3-70b-instruct API price — 4 gateways compared

from $0.10 per million input tokens · 1.4× spread between gateways · updated hourly

Gateway$ in /MTok$ out /MTok ContextInput price historyModel ID
OpenRouter fees$0.10$0.32131,072meta-llama/llama-3.3-70b-instruct
Glama$0.10$0.32128,000since 2026-08-20meta/llama-3.3-70b-instruct
Requesty$0.13$0.39128,000since 2026-08-20nebius/meta-llama/Llama-3.3-70B-Instruct
Novita$0.14$0.4012,288since 2026-08-20meta-llama/llama-3.3-70b-instruct

Fees beyond the token price

Changelog for llama-3.3-70b-instruct

llama-3.3-70b-instruct pricing FAQ

How much does llama-3.3-70b-instruct cost per million tokens?

The cheapest tracked gateway charges $0.10 per million input tokens and $0.32 per million output tokens (Glama). Prices refresh hourly.

Which gateway is cheapest for llama-3.3-70b-instruct?

Glama for input tokens and Glama for output tokens, among the 4 gateways we track.

Do llama-3.3-70b-instruct prices differ between gateways?

Yes — the most expensive tracked gateway charges 1.4× the cheapest for input tokens. Identical model, identical tokens, different bill.

← All LLM prices · API