Browse all Z.AI text models
Provider logo

GLM 5 Turbo

z-ai/glm-5-turbo
Back
Provider logo

GLM 5 Turbo

z-ai/glm-5-turbo
Back

Fast GLM 5 Turbo variant from Z-AI for general chat, coding, and tool use.

Added Mar 15, 2026

Context Window

202.8K

Max Output

131.1K

Avg output tokens (7d)

578 tokens

51%

Input Price (Auto)

$1.20/1M

Output Price (Auto)

$4.00/1M

Cache Read (Auto)

$0.24/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

Sourced from Artificial Analysis.

Intelligence Index

26.6

Better than 79% of models compared

Agentic work

T²-Bench Telecom (legacy)

Legacy fallback · Conversational AI agents in dual-control scenarios

98.5%

Better than 99% of models compared

Document reasoning

AA-LCR v1.1

Long context reasoning with updated grading

71.7%

Better than 66% of models compared

Reasoning

HLE

Humanity's Last Exam

27.8%

Better than 75% of models compared

IFBench

Instruction-following benchmark

73.2%

Better than 90% of models compared

CritPt

Research-level physics reasoning

0.3%

Coding

Terminal-Bench Hard (legacy)

Legacy fallback · Agentic coding and terminal use

33.3%

Better than 77% of models compared

Knowledge

AA-Omniscience Accuracy

Proportion of correctly answered questions

28.4%

AA-Omniscience Hallucination Rate

Rate of incorrect answers among non-correct responses

62.6%

Legacy benchmarks

GPQA Diamond (legacy)

Graduate-level scientific reasoning

84.7%

Better than 78% of models compared

AA-LCR (unversioned / legacy)

Long context reasoning evaluation

71.7%

Better than 66% of models compared

Last updated Oct 1, 2026

Artificial Analysis

Providers

Auto routing is available for this model. Explicit provider selection is not available.

Loading provider options…

Compare GLM 5 Turbo with similar models from the same provider or model family.

GLM 5V Turbo Thinking

z-ai/glm-5v-turbo:thinking

Thinking-enabled GLM 5V Turbo for image, video, and text inputs. Uses the same multimodal foundation model with more deliberate vision-grounded analysis, planning, and tool use.

GLM 5V Turbo

z-ai/glm-5v-turbo

Z.ai's native multimodal agent model for vision-based coding and agent workflows. This is the standard non-thinking variant for image, video, and text inputs, tuned for perceive-plan-execute loops, complex coding, and tool-driven task execution.

GLM 4.6 Turbo

z-ai/GLM-4.6-turbo

Fast variant of GLM 4.6 for general chat, coding, and analysis with improved latency and strong reasoning.

GLM 4.6 Turbo (Thinking)

z-ai/GLM-4.6-turbo:thinking

GLM 4.6 Turbo with thinking mode enabled for enhanced reasoning; shows internal reasoning and supports long context.

GLM 5.3 Flash Cybersecurity

z-ai/glm-5.3-flash-cybersecurity

GLM 5.3 Flash Cybersecurity is a cybersecurity-focused variant based on the uncensored model, with provider moderation for illegal activities. It supports always-on reasoning, image understanding, tool calling, and a 1,048,576-token context window.

GLM 5.3 Flash

z-ai/glm-5.3-flash

ox-alpha out of stealth! GLM-5.3 Flash is Z.ai's first natively multimodal GLM-5 model, with 320B total parameters and just 18B active parameters for efficient coding, agentic work, and precise 1M-token context. Its hybrid sparse-and-linear attention architecture helps it outperform GLM-5.2 at one-tenth the price while approaching Claude Opus 4.8 on coding and agentic benchmarks.