nvidia-nemotron-3-super-120b-a12b API price — 2 gateways compared
Cheapest right now: DeepInfra at $0.09/MTok in — 1.2× cheaper than the priciest gateway.
from $0.09 per
million input tokens · 1.2× spread between gateways
· updated continually
| Gateway | $ in /MTok | $ out /MTok |
Context | Input price history | Model ID | |
| DeepInfra | $0.09 | $0.40 | 262,144 | | nvidia/NVIDIA-Nemotron-3-Super-120B-A12B | use |
| Requesty | $0.10 | $0.50 | 262,144 | | deepinfra/nvidia/NVIDIA-Nemotron-3-Super-120B-A12B | use |
Changelog for nvidia-nemotron-3-super-120b-a12b
- requesty raised nvidia-nemotron-3-super-120b-a12b-output 11.1%: 0.45 → 0.5 usd_per_mtok
- requesty raised nvidia-nemotron-3-super-120b-a12b-input 11.1%: 0.09 → 0.1 usd_per_mtok
nvidia-nemotron-3-super-120b-a12b pricing FAQ
How much does nvidia-nemotron-3-super-120b-a12b cost per million tokens?
The cheapest tracked gateway charges $0.09 per million input tokens and $0.40 per million output tokens (DeepInfra). Prices refresh continually.
Which gateway is cheapest for nvidia-nemotron-3-super-120b-a12b?
DeepInfra for input tokens and DeepInfra for output tokens, among the 2 gateways we track.
Do nvidia-nemotron-3-super-120b-a12b prices differ between gateways?
Yes — the most expensive tracked gateway charges 1.2× the cheapest for input tokens. Identical model, identical tokens, different bill.
← All LLM prices · API