Z.AI · all models ›

ZGLM 5 TurboOPEN

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios.

203KContext window · tokens
$1.2Input price · per M tokens
$4Output price · per M tokens
Z.AIProvider

Prices update automatically — checked daily against provider list prices.

See model comparisons → Compare all model prices

Benchmarks & performance

Independent benchmark scores for GLM 5 Turbo, measured by Artificial Analysis. Higher is better.

Intelligence index39.1
GPQA84.7%
Humanity's Last Exam27.8%
Long Context Reasoning66.7%
SciCode43.6%
IFBench73.2%
τ²-Bench98.5%
Terminal-Bench Hard33.3%
vs. models measured heretop 28%
Scores higher than 72% of the 273 models measured by Artificial Analysis and tracked here.
Models at this level cost $1.13 per 1M tokens (median of 26) — this one costs $1.9.
Cheaper and better on this index: GLM 5.3 Flash · Gemini 3.7 Flash · Gemini 3.7 Flash (batch) and 29 more
Benchmark data by Artificial Analysis

About this model

GLM 5 Turbo is an open-weight AI model by Z.AI. You can download and self-host it for free; the prices below are hosted-API list prices, tracked daily, for when you prefer convenience over self-hosting.

Frequently asked questions

What is GLM 5 Turbo?

GLM 5 Turbo is an AI language model from Z.AI. It is open-weight: you can download it and run it on your own hardware, for free. It scores 39.1 on the Artificial Analysis intelligence index.

Is GLM 5 Turbo free?

The weights are free and open — you can self-host GLM 5 Turbo and pay nothing per token. If you prefer a hosted API, list prices are $1.2 per million input tokens and $4 per million output tokens.

What is GLM 5 Turbo good at?

Independent benchmarks from Artificial Analysis give it GPQA 84.7%, Humanity's Last Exam 27.8%, Long Context Reasoning 66.7%, SciCode 43.6%, IFBench 73.2%, τ²-Bench 98.5%, Terminal-Bench Hard 33.3%.

Can I self-host GLM 5 Turbo?

Yes. GLM 5 Turbo has open weights, so you can download it and run it on your own GPU or server with tools like Ollama, vLLM or llama.cpp — with no per-token cost.

Related models

GLM 5.2Z.AIGLM 5.3 FlashZ.AIGLM 5.3Z.AIGLM 5.1Z.AIGLM 4.7Z.AIGLM 5Z.AI