Browse all Z.AI text models
Provider logo

GLM 5

z-ai/glm-5
Back
Provider logo

GLM 5

z-ai/glm-5
Back

GLM-5 is a capable open-weight model from Zhipu with advanced reasoning and instruction following.

Added Feb 11, 2026

Model weights

Context Window

200.0K

Max Output

128.0K

Avg output tokens (7d)

269 tokens

22%

Input Price (Auto)

$0.60/1M

Output Price (Auto)

$1.90/1M

Cache Read (Auto)

$0.10/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

27.9

Better than 80% of models compared

Agentic work

T²-Bench Telecom (legacy)

Legacy fallback · Conversational AI agents in dual-control scenarios

98.2%

Better than 98% of models compared

Document reasoning

AA-LCR v1.1

Long context reasoning with updated grading

75.7%

Better than 74% of models compared

Reasoning

HLE

Humanity's Last Exam

29.3%

Better than 78% of models compared

IFBench

Instruction-following benchmark

72.3%

Better than 89% of models compared

CritPt

Research-level physics reasoning

2.0%

Coding

Terminal-Bench Hard (legacy)

Legacy fallback · Agentic coding and terminal use

43.2%

Better than 90% of models compared

Knowledge

AA-Omniscience Accuracy

Proportion of correctly answered questions

26.3%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

35.3%

Legacy benchmarks

GPQA Diamond (legacy)

Graduate-level scientific reasoning

82.0%

Better than 70% of models compared

AA-LCR (unversioned / legacy)

Long context reasoning evaluation

75.7%

Better than 74% of models compared

Last updated Oct 2, 2026

Artificial Analysis

Providers

Choose explicit providers for this model. Auto routing remains available as the default option.

Loading provider options…

Compare GLM 5 with similar models from the same provider or model family.

GLM 5.3 Flash Cybersecurity

z-ai/glm-5.3-flash-cybersecurity

GLM 5.3 Flash Cybersecurity is a cybersecurity-focused variant based on the uncensored model, with provider moderation for illegal activities. It supports always-on reasoning, image understanding, tool calling, and a 1,048,576-token context window.

GLM 5.3 Flash

z-ai/glm-5.3-flash

ox-alpha out of stealth! GLM-5.3 Flash is Z.ai's first natively multimodal GLM-5 model, with 320B total parameters and just 18B active parameters for efficient coding, agentic work, and precise 1M-token context. Its hybrid sparse-and-linear attention architecture helps it outperform GLM-5.2 at one-tenth the price while approaching Claude Opus 4.8 on coding and agentic benchmarks.

GLM 5.3

z-ai/glm-5.3

GLM-5.3 for long-horizon autonomous coding and engineering workflows. This variant defaults to the model's low reasoning tier for faster responses.

GLM 5.3 Thinking

z-ai/glm-5.3:thinking

GLM-5.3 with higher reasoning enabled for harder long-horizon coding, autonomous agent workflows, and complex engineering tasks.

GLM 5.3 Flash Uncensored

z-ai/glm-5.3-flash-uncensored

GLM 5.3 Flash Uncensored is an uncensored fine-tune of the efficient 320B mixture-of-experts reasoning model with provider-dependent vision support, built for unrestricted chat, creative writing, coding, agentic work, tool use, and long-context tasks.

GLM 5.2

z-ai/glm-5.2

GLM-5.2 is Z.AI's flagship model for long-horizon autonomous coding and engineering workflows. It is built to plan, execute, iterate, and optimize complex development tasks over extended runs. This variant keeps thinking disabled for faster direct responses.