gpudiff the public record of change in GPU cloud pricing
Follow the diffs: RSS · Free API · Sponsor this site

llama-4-scout-17b-16e-instruct API price — 2 gateways compared

from $0.10 per million input tokens · 1.8× spread between gateways · updated hourly

Gateway$ in /MTok$ out /MTok ContextInput price historyModel ID
DeepInfra$0.10$0.30327,680since 2026-08-20meta-llama/Llama-4-Scout-17B-16E-Instruct
Novita$0.18$0.59131,072since 2026-08-20meta-llama/llama-4-scout-17b-16e-instruct

Changelog for llama-4-scout-17b-16e-instruct

llama-4-scout-17b-16e-instruct pricing FAQ

How much does llama-4-scout-17b-16e-instruct cost per million tokens?

The cheapest tracked gateway charges $0.10 per million input tokens and $0.30 per million output tokens (DeepInfra). Prices refresh hourly.

Which gateway is cheapest for llama-4-scout-17b-16e-instruct?

DeepInfra for input tokens and DeepInfra for output tokens, among the 2 gateways we track.

Do llama-4-scout-17b-16e-instruct prices differ between gateways?

Yes — the most expensive tracked gateway charges 1.8× the cheapest for input tokens. Identical model, identical tokens, different bill.

← All LLM prices · API