All models

Llama 4 Maverick vs Qwen3.6 35B A3B

Llama 4 Maverick at $0.230 in and $0.920 out and Qwen3.6 35B A3B at $0.161 in and $1.15 out — per million tokens, from the same wallet.

M

Llama 4 Maverick

Meta

0.19¢

for a client proposal

$0.230 in · $0.920 out / 1M

1.05M context · released April 2025

Full page →
Q

Qwen3.6 35B A3B

Qwen

0.17¢

for a client proposal

$0.161 in · $1.15 out / 1M

262K context · released 3 months ago

Full page →

Asking all both at once — identical prompt, identical context — costs about 0.36¢ for a typical client proposal.

At a glance

  • Cheapest input: Qwen3.6 35B A3B at $0.161 per million tokens.
  • Cheapest output: Llama 4 Maverick at $0.920 per million tokens.
  • Largest context: Llama 4 Maverick at 1.05M tokens — about 1,500 pages.

Cost per task

Estimated totals at the rates above. Solid bar is input, lighter bar is output.

Quick code fix

2K in · 500 out
Llama 4 Maverick
0.09¢
Qwen3.6 35B A3B
0.09¢

Draft a client proposal from notes

5K in · 800 out
Llama 4 Maverick
0.19¢
Qwen3.6 35B A3B
0.17¢

Context-heavy session (Alyph workspace)

155K in · 4K out
Llama 4 Maverick
3.9¢
Qwen3.6 35B A3B

Analyze a 50-page legal contract

35K in · 1K out
Llama 4 Maverick
0.9¢
Qwen3.6 35B A3B
0.68¢

Large architecture refactor

250K in · 3K out
Llama 4 Maverick
$0.06
Qwen3.6 35B A3B
4.4¢

Solid part of each bar is input tokens, the lighter part is output. Models that don’t fit a task are marked instead of priced.

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free$0.18$0.370 tokens250k500k750k1000k
Llama 4 Maverick
Qwen3.6 35B A3B

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

Pricing

Llama 4 Maverick
Qwen3.6 35B A3B
Input / 1M tokens
$0.230
$0.161
Output / 1M tokens
$0.920
$1.15
Cached input / 1M
$0.058

Specs

Llama 4 Maverick
Qwen3.6 35B A3B
Provider
Released
April 2025
3 months ago
Context window
1.05M
262K
Max output
16K
262K
Input
Text and Images
Text, Images, and Video
Output
Text
Text
Reasoning
Answers directly
Shows its thinking
Knowledge cutoff
2024-08-31

Try it with real numbers

2K tokens in · 500 tokens out.

Llama 4 Maverick

Meta

0.09¢

for this task

50% input · 50% output

Full pricing →

Qwen3.6 35B A3B

Qwen

0.09¢

for this task

36% input · 64% output

Full pricing →

All 2, same prompt and context: 0.18¢. That is the entire cost of the comparison.

Your $5 welcome credit covers about 2,751 of these.

Which should you pick?

For long documents, Llama 4 Maverick has the largest context window here — 1.05M tokens.

On price, Qwen3.6 35B A3B is the cheapest of the two on a typical task, at about 0.17¢.

If you want to watch the model think, Qwen3.6 35B A3B shows reasoning; the other answers directly.

Or don’t pick. Send the identical prompt to all both, read the answers side by side, and keep the winner. How to run a fair bake-off →

Questions, answered

Which is cheaper, Llama 4 Maverick or Qwen3.6 35B A3B?

On a typical client proposal (5K tokens in, 800 out): Qwen3.6 35B A3B at 0.17¢; Llama 4 Maverick at 0.19¢. Long prompts can change the order when long-context rates apply.

Which has the largest context window?

Llama 4 Maverick — 1.05M tokens, about 1,500 pages.

Can I run Llama 4 Maverick and Qwen3.6 35B A3B side by side on Alyph?

Yes — that is what Alyph is built for: the identical prompt with identical context to all 2, answers rendered side by side, billed from one wallet. This exact combination costs about 0.36¢ for a typical task.

Ask both at once

Identical prompt, identical context, answers side by side — about 0.36¢ for a typical client proposal. Free to start, with $5 of credit.