gpudiff the public record of change in GPU cloud pricing
Follow the diffs: RSS · Free API · Sponsor this site

llama-3.3-70b-instruct-turbo API price — 2 gateways compared

Cheapest right now: DeepInfra at $0.10/MTok in — 1.2× cheaper than the priciest gateway.

from $0.10 per million input tokens · 1.2× spread between gateways · updated continually

Gateway$ in /MTok$ out /MTok ContextInput price historyModel ID
DeepInfra$0.10$0.32131,072meta-llama/Llama-3.3-70B-Instruct-Turbouse
Requesty$0.12$0.30131,072deepinfra/meta-llama/Llama-3.3-70B-Instruct-Turbouse

Changelog for llama-3.3-70b-instruct-turbo

llama-3.3-70b-instruct-turbo pricing FAQ

How much does llama-3.3-70b-instruct-turbo cost per million tokens?

The cheapest tracked gateway charges $0.10 per million input tokens and $0.30 per million output tokens (DeepInfra). Prices refresh continually.

Which gateway is cheapest for llama-3.3-70b-instruct-turbo?

DeepInfra for input tokens and Requesty for output tokens, among the 2 gateways we track.

Do llama-3.3-70b-instruct-turbo prices differ between gateways?

Yes — the most expensive tracked gateway charges 1.2× the cheapest for input tokens. Identical model, identical tokens, different bill.

← All LLM prices · API