All models

GPT-3.5 Turbo (older v0613) vs GLM 5

GPT-3.5 Turbo (older v0613) at $1.15 in and $2.30 out and GLM 5 at $1.09 in and $2.93 out — per million tokens, from the same wallet.

GPT-3.5 Turbo (older v0613)

OpenAI

0.76¢

for a client proposal

$1.15 in · $2.30 out / 1M

4K context · released January 2024

Full page →
Z

GLM 5

Z.ai

0.78¢

for a client proposal

$1.09 in · $2.93 out / 1M

205K context · released 6 months ago

Full page →

Asking all both at once — identical prompt, identical context — costs about 1.5¢ for a typical client proposal.

At a glance

  • Cheapest input: GLM 5 at $1.09 per million tokens.
  • Cheapest output: GPT-3.5 Turbo (older v0613) at $2.30 per million tokens.
  • Largest context: GLM 5 at 205K tokens — about 290 pages.

Cost per task

Estimated totals at the rates above. Solid bar is input, lighter bar is output.

Quick code fix

2K in · 500 out
GPT-3.5 Turbo (older v0613)
0.34¢
GLM 5
0.37¢

Draft a client proposal from notes

5K in · 800 out
GPT-3.5 Turbo (older v0613)
needs 5K ctx
GLM 5
0.78¢

Context-heavy session (Alyph workspace)

155K in · 4K out
GPT-3.5 Turbo (older v0613)
needs 155K ctx
GLM 5
$0.18

Analyze a 50-page legal contract

35K in · 1K out
GPT-3.5 Turbo (older v0613)
needs 35K ctx
GLM 5
4.1¢

Large architecture refactor

250K in · 3K out
GPT-3.5 Turbo (older v0613)
needs 250K ctx
GLM 5
needs 250K ctx

Solid part of each bar is input tokens, the lighter part is output. Models that don’t fit a task are marked instead of priced.

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free$0.73$1.460 tokens250k500k750k1000k
GPT-3.5 Turbo (older v0613)
GLM 5

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

Pricing

GPT-3.5 Turbo (older v0613)
GLM 5
Input / 1M tokens
$1.15
$1.09
Output / 1M tokens
$2.30
$2.93
Cached input / 1M
$0.230

Specs

GPT-3.5 Turbo (older v0613)
GLM 5
Provider
Released
January 2024
6 months ago
Context window
4K
205K
Max output
4K
131K
Input
Text
Text
Output
Text
Text
Reasoning
Answers directly
Shows its thinking
Knowledge cutoff
2021-09-30

Try it with real numbers

2K tokens in · 500 tokens out.

GPT-3.5 Turbo (older v0613)

OpenAI

0.34¢

for this task

67% input · 33% output

Full pricing →

GLM 5

Z.ai

0.37¢

for this task

60% input · 40% output

Full pricing →

All 2, same prompt and context: 0.71¢. That is the entire cost of the comparison.

Your $5 welcome credit covers about 704 of these.

Which should you pick?

For long documents, GLM 5 has the largest context window here — 205K tokens.

On price, GPT-3.5 Turbo (older v0613) is the cheapest of the two on a typical task, at about 0.76¢.

If you want to watch the model think, GLM 5 shows reasoning; the other answers directly.

Or don’t pick. Send the identical prompt to all both, read the answers side by side, and keep the winner. How to run a fair bake-off →

Questions, answered

Which is cheaper, GPT-3.5 Turbo (older v0613) or GLM 5?

On a typical client proposal (5K tokens in, 800 out): GPT-3.5 Turbo (older v0613) at 0.76¢; GLM 5 at 0.78¢. Long prompts can change the order when long-context rates apply.

Which has the largest context window?

GLM 5 — 205K tokens, about 290 pages.

Can I run GPT-3.5 Turbo (older v0613) and GLM 5 side by side on Alyph?

Yes — that is what Alyph is built for: the identical prompt with identical context to all 2, answers rendered side by side, billed from one wallet. This exact combination costs about 1.5¢ for a typical task.

Ask both at once

Identical prompt, identical context, answers side by side — about 1.5¢ for a typical client proposal. Free to start, with $5 of credit.