All models

Qwen3 8B vs MiMo-V2.5

Qwen3 8B at $0.135 in and $0.523 out and MiMo-V2.5 at $0.161 in and $0.322 out — per million tokens, from the same wallet.

Q

Qwen3 8B

Qwen

0.11¢

for a client proposal

$0.135 in · $0.523 out / 1M

131K context · released April 2025

Full page →
X

MiMo-V2.5

Xiaomi

0.11¢

for a client proposal

$0.161 in · $0.322 out / 1M

1.05M context · released 4 months ago

Full page →

Asking all both at once — identical prompt, identical context — costs about 0.22¢ for a typical client proposal.

At a glance

  • Cheapest input: Qwen3 8B at $0.135 per million tokens.
  • Cheapest output: MiMo-V2.5 at $0.322 per million tokens.
  • Largest context: MiMo-V2.5 at 1.05M tokens — about 1,500 pages.

Cost per task

Estimated totals at the rates above. Solid bar is input, lighter bar is output.

Quick code fix

2K in · 500 out
Qwen3 8B
0.05¢
MiMo-V2.5
0.05¢

Draft a client proposal from notes

5K in · 800 out
Qwen3 8B
0.11¢
MiMo-V2.5
0.11¢

Context-heavy session (Alyph workspace)

155K in · 4K out
Qwen3 8B
needs 155K ctx
MiMo-V2.5
2.6¢

Analyze a 50-page legal contract

35K in · 1K out
Qwen3 8B
0.52¢
MiMo-V2.5
0.6¢

Large architecture refactor

250K in · 3K out
Qwen3 8B
needs 250K ctx
MiMo-V2.5
4.1¢

Solid part of each bar is input tokens, the lighter part is output. Models that don’t fit a task are marked instead of priced.

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free$0.11$0.210 tokens250k500k750k1000k
Qwen3 8B
MiMo-V2.5

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

Pricing

Qwen3 8B
MiMo-V2.5
Input / 1M tokens
$0.135
$0.161
Output / 1M tokens
$0.523
$0.322
Cached input / 1M
$0.0032

Specs

Qwen3 8B
MiMo-V2.5
Provider
Released
April 2025
4 months ago
Context window
131K
1.05M
Max output
8K
131K
Input
Text
Text, Audio, Images, and Video
Output
Text
Text
Reasoning
Shows its thinking
Shows its thinking
Knowledge cutoff
2025-03-31

Try it with real numbers

2K tokens in · 500 tokens out.

Qwen3 8B

Qwen

0.05¢

for this task

51% input · 49% output

Full pricing →

MiMo-V2.5

Xiaomi

0.05¢

for this task

67% input · 33% output

Full pricing →

All 2, same prompt and context: 0.1¢. That is the entire cost of the comparison.

Your $5 welcome credit covers about 4,932 of these.

Which should you pick?

For long documents, MiMo-V2.5 has the largest context window here — 1.05M tokens.

On price, MiMo-V2.5 is the cheapest of the two on a typical task, at about 0.11¢.

For images or files, MiMo-V2.5 can read them directly; Qwen3 8B is text-only.

Or don’t pick. Send the identical prompt to all both, read the answers side by side, and keep the winner. How to run a fair bake-off →

Questions, answered

Which is cheaper, Qwen3 8B or MiMo-V2.5?

On a typical client proposal (5K tokens in, 800 out): MiMo-V2.5 at 0.11¢; Qwen3 8B at 0.11¢. Long prompts can change the order when long-context rates apply.

Which has the largest context window?

MiMo-V2.5 — 1.05M tokens, about 1,500 pages.

Can I run Qwen3 8B and MiMo-V2.5 side by side on Alyph?

Yes — that is what Alyph is built for: the identical prompt with identical context to all 2, answers rendered side by side, billed from one wallet. This exact combination costs about 0.22¢ for a typical task.

Ask both at once

Identical prompt, identical context, answers side by side — about 0.22¢ for a typical client proposal. Free to start, with $5 of credit.