All Z.ai models
Z

GLM 4.7 Flash

Z.ai · released 7 months ago

GLM 4.7 Flash is a Z.ai model with a 203K token context window — about 290 pages of text. It reads Text and replies in Text. It can show its reasoning before it answers.

$0.069

input / 1M tokens

$0.460

output / 1M tokens

203K

context — about 290 pages

0.07¢

a client proposal

Pricing

Input
$0.069 per million tokens
Output
$0.460 per million tokens
Cached input
$0.011 per million tokens

Billed from your Alyph wallet. Prices can change when the provider changes theirs — last checked August 9, 2026.

The numbers

Provider
Released
7 months ago (2026-01-19)
Context window
203K tokens — about 290 pages
Max output
16K tokens per reply
Input
Text
Output
Text
Reasoning
Shows its thinking before answering

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free$0.07$0.150 tokens250k500k750k1000k
GLM 4.7 Flash

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

What a task costs

2K tokens in · 500 tokens out.

GLM 4.7 Flash

Z.ai

0.04¢

for this task

38% input · 62% output

Full pricing →

Your $5 welcome credit covers about 13,586 of these.

Compare GLM 4.7 Flash

Closest in price on a typical task:

Questions, answered

How much does GLM 4.7 Flash cost on Alyph?

GLM 4.7 Flash costs $0.069 per million input tokens and $0.460 per million output tokens on Alyph. A typical task — 5K tokens in, 800 out — works out to about 0.07¢. Billed from your Alyph wallet, with hard spending limits. Prices can change when the provider changes theirs; this page was last checked on August 9, 2026.

What is GLM 4.7 Flash's context window?

203K tokens — about 290 pages. Replies can be up to 16K tokens long.

Can GLM 4.7 Flash read images and files?

GLM 4.7 Flash is text-only. The canvas still serializes uploads into plain text, so you can work with files — but it cannot see images.

Does GLM 4.7 Flash show its reasoning?

Yes. GLM 4.7 Flash can expose its thinking before the final answer, so you can watch how it got there.

Test GLM 4.7 Flash against Qwen3.5-9B

One prompt, both models at the same time, the better answer wins. Free to start, with $5 of credit.