All Z.ai models
Z

GLM 4.6

Z.ai · released September 2025

GLM 4.6 is a Z.ai model with a 203K token context window — about 290 pages of text. It reads Text and replies in Text. It can show its reasoning before it answers.

$0.575

input / 1M tokens

$2.30

output / 1M tokens

203K

context — about 290 pages

0.47¢

a client proposal

Pricing

Input
$0.575 per million tokens
Output
$2.30 per million tokens
Cached input
$0.115 per million tokens

Billed from your Alyph wallet. Prices can change when the provider changes theirs — last checked August 9, 2026.

The numbers

Provider
Released
September 2025 (2025-09-30)
Context window
203K tokens — about 290 pages
Max output
131K tokens per reply
Input
Text
Output
Text
Reasoning
Shows its thinking before answering
Knowledge cutoff
2025-03-31

Cost vs Token Amount

Drag the slider to adjust the split between input and output workload.

80% Input20% Output
free$0.46$0.920 tokens250k500k750k1000k
GLM 4.6

Adjust the ratio slider to change how output tokens influence the final price for equivalent workloads.

What a task costs

2K tokens in · 500 tokens out.

GLM 4.6

Z.ai

0.23¢

for this task

50% input · 50% output

Full pricing →

Your $5 welcome credit covers about 2,173 of these.

Compare GLM 4.6

Closest in price on a typical task:

Questions, answered

How much does GLM 4.6 cost on Alyph?

GLM 4.6 costs $0.575 per million input tokens and $2.30 per million output tokens on Alyph. A typical task — 5K tokens in, 800 out — works out to about 0.47¢. Billed from your Alyph wallet, with hard spending limits. Prices can change when the provider changes theirs; this page was last checked on August 9, 2026.

What is GLM 4.6's context window?

203K tokens — about 290 pages. Replies can be up to 131K tokens long.

Can GLM 4.6 read images and files?

GLM 4.6 is text-only. The canvas still serializes uploads into plain text, so you can work with files — but it cannot see images.

Does GLM 4.6 show its reasoning?

Yes. GLM 4.6 can expose its thinking before the final answer, so you can watch how it got there.

Test GLM 4.6 against Qwen2.5 Coder 32B Instruct

One prompt, both models at the same time, the better answer wins. Free to start, with $5 of credit.