All models

Gemma 2 27B vs Qwen3.5 397B A17B

Gemma 2 27B at $0.748 in and $0.748 out and Qwen3.5 397B A17B at $0.449 in and $2.69 out — per million tokens, from the same wallet.

Gemma 2 27B

Google

0.43¢

for a client proposal

$0.748 in · $0.748 out / 1M

8K context · released July 2024

Full page →
Q

Qwen3.5 397B A17B

Qwen

0.44¢

for a client proposal

$0.449 in · $2.69 out / 1M

262K context · released 6 months ago

Full page →

Asking all both at once — identical prompt, identical context — costs about 0.87¢ for a typical client proposal.

At a glance

  • Cheapest input: Qwen3.5 397B A17B at $0.449 per million tokens.
  • Cheapest output: Gemma 2 27B at $0.748 per million tokens.
  • Largest context: Qwen3.5 397B A17B at 262K tokens — about 370 pages.

Cost per task

Estimated totals at the rates above. Solid bar is input, lighter bar is output.

Quick code fix

2K in · 500 out
Gemma 2 27B
0.19¢
Qwen3.5 397B A17B
0.22¢

Draft a client proposal from notes

5K in · 800 out
Gemma 2 27B
0.43¢
Qwen3.5 397B A17B
0.44¢

Context-heavy session (Alyph workspace)

155K in · 4K out
Gemma 2 27B
needs 155K ctx
Qwen3.5 397B A17B
$0.08

Analyze a 50-page legal contract

35K in · 1K out
Gemma 2 27B
needs 35K ctx
Qwen3.5 397B A17B
1.8¢

Large architecture refactor

250K in · 3K out
Gemma 2 27B
needs 250K ctx
Qwen3.5 397B A17B
$0.12

Solid part of each bar is input tokens, the lighter part is output. Models that don’t fit a task are marked instead of priced.

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free$0.45$0.900 tokens250k500k750k1000k
Gemma 2 27B
Qwen3.5 397B A17B

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

Pricing

Gemma 2 27B
Qwen3.5 397B A17B
Input / 1M tokens
$0.748
$0.449
Output / 1M tokens
$0.748
$2.69

Specs

Gemma 2 27B
Qwen3.5 397B A17B
Provider
Released
July 2024
6 months ago
Context window
8K
262K
Max output
2K
66K
Input
Text
Text, Images, and Video
Output
Text
Text
Reasoning
Answers directly
Shows its thinking
Knowledge cutoff
2024-06-30

Try it with real numbers

2K tokens in · 500 tokens out.

Gemma 2 27B

Google

0.19¢

for this task

80% input · 20% output

Full pricing →

Qwen3.5 397B A17B

Qwen

0.22¢

for this task

40% input · 60% output

Full pricing →

All 2, same prompt and context: 0.41¢. That is the entire cost of the comparison.

Your $5 welcome credit covers about 1,216 of these.

Which should you pick?

For long documents, Qwen3.5 397B A17B has the largest context window here — 262K tokens.

On price, Gemma 2 27B is the cheapest of the two on a typical task, at about 0.43¢.

For images or files, Qwen3.5 397B A17B can read them directly; Gemma 2 27B is text-only.

If you want to watch the model think, Qwen3.5 397B A17B shows reasoning; the other answers directly.

Or don’t pick. Send the identical prompt to all both, read the answers side by side, and keep the winner. How to run a fair bake-off →

Questions, answered

Which is cheaper, Gemma 2 27B or Qwen3.5 397B A17B?

On a typical client proposal (5K tokens in, 800 out): Gemma 2 27B at 0.43¢; Qwen3.5 397B A17B at 0.44¢. Long prompts can change the order when long-context rates apply.

Which has the largest context window?

Qwen3.5 397B A17B — 262K tokens, about 370 pages.

Can I run Gemma 2 27B and Qwen3.5 397B A17B side by side on Alyph?

Yes — that is what Alyph is built for: the identical prompt with identical context to all 2, answers rendered side by side, billed from one wallet. This exact combination costs about 0.87¢ for a typical task.

Ask both at once

Identical prompt, identical context, answers side by side — about 0.87¢ for a typical client proposal. Free to start, with $5 of credit.