gpudiff the public record of change in GPU cloud pricing
Follow the diffs: RSS · Free API · Sponsor this site

llama-4-maverick-17b-128e-instruct-fp8 API price — 2 gateways compared

Cheapest right now: Requesty at $0.20/MTok in — 1.4× cheaper than the priciest gateway.

from $0.20 per million input tokens · 1.4× spread between gateways · updated continually

Gateway$ in /MTok$ out /MTok ContextInput price historyModel ID
Requesty$0.20$0.851,048,576novita/meta-llama/llama-4-maverick-17b-128e-instruct-fp8use
Novita$0.27$0.851,048,576meta-llama/llama-4-maverick-17b-128e-instruct-fp8use

Changelog for llama-4-maverick-17b-128e-instruct-fp8

llama-4-maverick-17b-128e-instruct-fp8 pricing FAQ

How much does llama-4-maverick-17b-128e-instruct-fp8 cost per million tokens?

The cheapest tracked gateway charges $0.20 per million input tokens and $0.85 per million output tokens (Requesty). Prices refresh continually.

Which gateway is cheapest for llama-4-maverick-17b-128e-instruct-fp8?

Requesty for input tokens and Novita for output tokens, among the 2 gateways we track.

Do llama-4-maverick-17b-128e-instruct-fp8 prices differ between gateways?

Yes — the most expensive tracked gateway charges 1.4× the cheapest for input tokens. Identical model, identical tokens, different bill.

← All LLM prices · API