All models

Gemma 3 4B vs Mistral Small 3

Gemma 3 4B at $0.058 in and $0.115 out and Mistral Small 3 at $0.058 in and $0.092 out — per million tokens, from the same wallet.

Gemma 3 4B

Google

0.04¢

for a client proposal

$0.058 in · $0.115 out / 1M

131K context · released March 2025

Full page →
M

Mistral Small 3

Mistral AI

0.04¢

for a client proposal

$0.058 in · $0.092 out / 1M

33K context · released January 2025

Full page →

Asking all both at once — identical prompt, identical context — costs about 0.07¢ for a typical client proposal.

At a glance

  • All 2 cost the same on input.
  • Cheapest output: Mistral Small 3 at $0.092 per million tokens.
  • Largest context: Gemma 3 4B at 131K tokens — about 190 pages.

Cost per task

Estimated totals at the rates above. Solid bar is input, lighter bar is output.

Quick code fix

2K in · 500 out
Gemma 3 4B
0.02¢
Mistral Small 3
0.02¢

Draft a client proposal from notes

5K in · 800 out
Gemma 3 4B
0.04¢
Mistral Small 3
0.04¢

Context-heavy session (Alyph workspace)

155K in · 4K out
Gemma 3 4B
needs 155K ctx
Mistral Small 3
needs 155K ctx

Analyze a 50-page legal contract

35K in · 1K out
Gemma 3 4B
0.21¢
Mistral Small 3
needs 35K ctx

Large architecture refactor

250K in · 3K out
Gemma 3 4B
needs 250K ctx
Mistral Small 3
needs 250K ctx

Solid part of each bar is input tokens, the lighter part is output. Models that don’t fit a task are marked instead of priced.

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free3.5¢$0.070 tokens250k500k750k1000k
Gemma 3 4B
Mistral Small 3

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

Pricing

Gemma 3 4B
Mistral Small 3
Input / 1M tokens
$0.058
$0.058
Output / 1M tokens
$0.115
$0.092

Specs

Gemma 3 4B
Mistral Small 3
Released
March 2025
January 2025
Context window
131K
33K
Max output
16K
16K
Input
Text and Images
Text
Output
Text
Text
Reasoning
Answers directly
Answers directly
Knowledge cutoff
2024-08-31
2023-10-31

Try it with real numbers

2K tokens in · 500 tokens out.

Gemma 3 4B

Google

0.02¢

for this task

67% input · 33% output

Full pricing →

Mistral Small 3

Mistral AI

0.02¢

for this task

71% input · 29% output

Full pricing →

All 2, same prompt and context: 0.03¢. That is the entire cost of the comparison.

Your $5 welcome credit covers about 14,992 of these.

Which should you pick?

For long documents, Gemma 3 4B has the largest context window here — 131K tokens.

On price, Mistral Small 3 is the cheapest of the two on a typical task, at about 0.04¢.

For images or files, Gemma 3 4B can read them directly; Mistral Small 3 is text-only.

Or don’t pick. Send the identical prompt to all both, read the answers side by side, and keep the winner. How to run a fair bake-off →

Questions, answered

Which is cheaper, Gemma 3 4B or Mistral Small 3?

On a typical client proposal (5K tokens in, 800 out): Mistral Small 3 at 0.04¢; Gemma 3 4B at 0.04¢. Long prompts can change the order when long-context rates apply.

Which has the largest context window?

Gemma 3 4B — 131K tokens, about 190 pages.

Can I run Gemma 3 4B and Mistral Small 3 side by side on Alyph?

Yes — that is what Alyph is built for: the identical prompt with identical context to all 2, answers rendered side by side, billed from one wallet. This exact combination costs about 0.07¢ for a typical task.

Ask both at once

Identical prompt, identical context, answers side by side — about 0.07¢ for a typical client proposal. Free to start, with $5 of credit.