All models

Nemotron 3 Ultra vs Qwen3 Coder Plus

Nemotron 3 Ultra at $0.690 in and $4.14 out and Qwen3 Coder Plus at $0.748 in and $3.74 out — per million tokens, from the same wallet.

N

Nemotron 3 Ultra

NVIDIA

0.68¢

for a client proposal

$0.690 in · $4.14 out / 1M

512K context · released 2 months ago

Full page →
Q

Qwen3 Coder Plus

Qwen

0.67¢

for a client proposal

$0.748 in · $3.74 out / 1M

1M context · released September 2025

Full page →

Asking all both at once — identical prompt, identical context — costs about 1.3¢ for a typical client proposal.

At a glance

  • Cheapest input: Nemotron 3 Ultra at $0.690 per million tokens.
  • Cheapest output: Qwen3 Coder Plus at $3.74 per million tokens.
  • Largest context: Qwen3 Coder Plus at 1M tokens — about 1,400 pages.

Cost per task

Estimated totals at the rates above. Solid bar is input, lighter bar is output.

Quick code fix

2K in · 500 out
Nemotron 3 Ultra
0.34¢
Qwen3 Coder Plus
0.34¢

Draft a client proposal from notes

5K in · 800 out
Nemotron 3 Ultra
0.68¢
Qwen3 Coder Plus
0.67¢

Context-heavy session (Alyph workspace)

155K in · 4K out
Nemotron 3 Ultra
$0.12
Qwen3 Coder Plus
$0.39

Long-context rate applied at this input size.

Analyze a 50-page legal contract

35K in · 1K out
Nemotron 3 Ultra
2.8¢
Qwen3 Coder Plus
$0.05

Long-context rate applied at this input size.

Large architecture refactor

250K in · 3K out
Nemotron 3 Ultra
$0.18
Qwen3 Coder Plus
$0.59

Long-context rate applied at this input size.

Solid part of each bar is input tokens, the lighter part is output. Models that don’t fit a task are marked instead of priced.

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free$2.02$4.040 tokens250k500k750k1000k
Nemotron 3 Ultra
Qwen3 Coder Plus

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

Pricing

Nemotron 3 Ultra
Qwen3 Coder Plus
Input / 1M tokens
$0.690
$0.748
Output / 1M tokens
$4.14
$3.74
Cached input / 1M
$0.230
$0.149

Specs

Nemotron 3 Ultra
Qwen3 Coder Plus
Provider
Released
2 months ago
September 2025
Context window
512K
1M
Max output
66K
Input
Text
Text
Output
Text
Text
Reasoning
Shows its thinking
Answers directly
Knowledge cutoff
2025-06-30

Try it with real numbers

2K tokens in · 500 tokens out.

Nemotron 3 Ultra

NVIDIA

0.34¢

for this task

40% input · 60% output

Full pricing →

Qwen3 Coder Plus

Qwen

0.34¢

for this task

44% input · 56% output

Full pricing →

All 2, same prompt and context: 0.68¢. That is the entire cost of the comparison.

Your $5 welcome credit covers about 733 of these.

Which should you pick?

For long documents, Qwen3 Coder Plus has the largest context window here — 1M tokens.

On price, Qwen3 Coder Plus is the cheapest of the two on a typical task, at about 0.67¢.

If you want to watch the model think, Nemotron 3 Ultra shows reasoning; the other answers directly.

Or don’t pick. Send the identical prompt to all both, read the answers side by side, and keep the winner. How to run a fair bake-off →

Questions, answered

Which is cheaper, Nemotron 3 Ultra or Qwen3 Coder Plus?

On a typical client proposal (5K tokens in, 800 out): Qwen3 Coder Plus at 0.67¢; Nemotron 3 Ultra at 0.68¢. Long prompts can change the order when long-context rates apply.

Which has the largest context window?

Qwen3 Coder Plus — 1M tokens, about 1,400 pages.

Can I run Nemotron 3 Ultra and Qwen3 Coder Plus side by side on Alyph?

Yes — that is what Alyph is built for: the identical prompt with identical context to all 2, answers rendered side by side, billed from one wallet. This exact combination costs about 1.3¢ for a typical task.

Ask both at once

Identical prompt, identical context, answers side by side — about 1.3¢ for a typical client proposal. Free to start, with $5 of credit.