gpudiff the public record of change in GPU cloud pricing
Follow the diffs: RSS · Free API · Sponsor this site

qwen3-vl-8b-instruct API price — 3 gateways compared

from $0.08 per million input tokens · 1.5× spread between gateways · updated hourly

Gateway$ in /MTok$ out /MTok ContextInput price historyModel ID
Novita$0.08$0.50131,072since 2026-08-20qwen/qwen3-vl-8b-instruct
OpenRouter fees$0.12$0.46262,144qwen/qwen3-vl-8b-instruct
Glama$0.12$0.46131,072since 2026-08-20alibaba/qwen3-vl-8b-instruct

Fees beyond the token price

Changelog for qwen3-vl-8b-instruct

qwen3-vl-8b-instruct pricing FAQ

How much does qwen3-vl-8b-instruct cost per million tokens?

The cheapest tracked gateway charges $0.08 per million input tokens and $0.46 per million output tokens (Novita). Prices refresh hourly.

Which gateway is cheapest for qwen3-vl-8b-instruct?

Novita for input tokens and Glama for output tokens, among the 3 gateways we track.

Do qwen3-vl-8b-instruct prices differ between gateways?

Yes — the most expensive tracked gateway charges 1.5× the cheapest for input tokens. Identical model, identical tokens, different bill.

← All LLM prices · API