GLM-4.6, Zhipu's flagship text model with 256K context window and advanced reasoning capabilities. Direct via Z-AI (Zhipu).
Added Dec 11, 2025
Context Window
256.0K
Max Output
65.5K
Avg output tokens (7d)
942 tokens
Input Price (Auto)
$0.50/1M
Output Price (Auto)
$2.00/1M
Cache Read (Auto)
$0.100/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
14.9
Agentic work
T²-Bench Telecom (legacy)
Legacy fallback · Conversational AI agents in dual-control scenarios
76.9%
Better than 68% of models compared
Document reasoning
AA-LCR v1.1
Long context reasoning with updated grading
26.3%
Better than 27% of models compared
Reasoning
HLE
Humanity's Last Exam
5.5%
Better than 33% of models compared
IFBench
Instruction-following benchmark
36.7%
Better than 30% of models compared
CritPt
Research-level physics reasoning
0.0%
Coding
Terminal-Bench Hard (legacy)
Legacy fallback · Agentic coding and terminal use
28.8%
Better than 71% of models compared
LiveCodeBench
Contamination-free coding benchmark
56.1%
Better than 63% of models compared
Math
AIME 2025
American Invitational Mathematics Examination 2025
44.3%
Better than 45% of models compared
Knowledge
MMLU-Pro
Professional and academic subject knowledge
78.4%
Better than 60% of models compared
AA-Omniscience Accuracy
Proportion of correctly answered questions
21.4%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
67.6%
Legacy benchmarks
GPQA Diamond (legacy)
Graduate-level scientific reasoning
63.2%
Better than 38% of models compared
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
26.3%
Better than 27% of models compared
Last updated Oct 2, 2026
Artificial AnalysisProviders
Choose explicit providers for this model. Auto routing remains available as the default option.
Loading provider options…
Related text models
Compare GLM 4.6 Original with similar models from the same provider or model family.
GLM 5 Original
z-ai/glm-5-originalGLM-5 is a capable open-weight model from Zhipu with advanced reasoning and instruction following. This Original variant uses the model maker's direct endpoint.
GLM 5 Original Thinking
z-ai/glm-5-original:thinkingGLM-5 original with extended thinking capabilities for complex reasoning.
GLM 4.7 Original
z-ai/glm-4.7-originalGLM-4.7 is a next-gen GLM series text model with stronger reasoning, long-context chat, and reliable tool use. Routed directly via Z-AI (Zhipu).
GLM 4.7 Original Thinking
z-ai/glm-4.7-original:thinkingGLM-4.7 original with extended thinking capabilities for complex reasoning.
GLM 4.6V Original
z-ai/glm-4.6v-originalGLM-4.6V scales its context window to 128k tokens in training, and achieves SoTA performance in visual understanding among models of similar parameter scales. Integrates native Function Calling capabilities, bridging 'visual perception' and 'executable action' for multimodal agents. Direct via Z-AI (Zhipu).
GLM 5.3 Flash Cybersecurity
z-ai/glm-5.3-flash-cybersecurityGLM 5.3 Flash Cybersecurity is a cybersecurity-focused variant based on the uncensored model, with provider moderation for illegal activities. It supports always-on reasoning, image understanding, tool calling, and a 1,048,576-token context window.