Qwen3 235B API prices across providers

Identical open weights, different invoices. All prices USD per 1M tokens, verified against provider sources. Last full sweep:

Qwen3 235B hosted inference prices by provider, USD per 1M tokens
Provider In $/M Out $/M Spread vs cheapest Context tok/s Last verified Go link
Together 0.2 0.6 1.00× cheapest 262k 40 · manual-csv Go ↗

Cache-hit pricing excluded — model your real mix in the Cost Calculator (ships week 2). Spread = output price ÷ cheapest output price.

Measured floor: Together at $0.6/M out. Widest spread on this model today: 1.00×.

What a month costs

Worked examples at this model's verified prices (2026-08-23 sweep); cache-hit discounts excluded.

Monthly USD cost examples for Qwen3 235B by provider
Monthly workloadNovita AIDeepInfraTogetherOpenRouter
Light — 1M in / 0.25M out $0.72$0.64$0.73$0.79
Mid — 10M in / 2M out $6.48$5.8$6.6$7.16
Heavy — 50M in / 10M out $32.4$29$33$35.8

Head-to-heads: DeepSeek V3.x · DeepSeek R1 · MiniMax M3

Also tracked: DeepSeek V3.x · DeepSeek R1 · Qwen3 Coder · Llama 4 Maverick · Kimi K2 · GLM 4.7 · GPT-OSS 120B · Gemma 3 27B · MiniMax M3

How we verify

  1. Pull — provider pricing pages and /models endpoints, daily at 06:00 UTC.
  2. Probe — fixed 1k-in/256-out prompt per provider; billed tokens and latency recorded.
  3. Diff — changes beyond ±1% raise a price move; every row carries its own last-verified date, never a site-wide stamp.
  4. Disclose — affiliate Go links are labeled and never reorder this table; methodology and raw probes are published.
Freshness policy: rows older than 14 days carry the amber stale flag until re-probed. Stale data beats fresh guesses.

Spot a wrong number? File a correction — we re-probe within one sweep cycle.