GLM-5 is the base model of Zhipu’s GLM-5 generation on z.ai — the generation’s workhorse, priced below the GLM-5.2 flagship. It is an OpenAI-compatible reasoning chat model with tool and JSON support.
z.ai publishes GLM-5 at $1.00 per 1M input tokens and $3.20 per 1M output tokens — cheaper than the GLM-5.2 flagship while staying in the same generation.
We registered it with a large (~200K-token) context window as a provisional figure; GLM-5-generation documentation is thin, so confirm the exact window on z.ai before relying on it.
Architecture & training signals
GLM-5 is the base of Zhipu AI’s GLM-5 generation. Public technical detail is still limited as of July 2026; we treat it as a large reasoning-oriented chat model with tool-calling and structured JSON over an OpenAI-compatible endpoint. Like the rest of the GLM line, it returns a non-standard reasoning_content field alongside content in its OpenAI-compatible responses; integrations should read content for the final answer and treat reasoning_content as an optional trace.
Where it shines
- A balanced quality-vs-price point within the newest GLM generation.
- Long-context reasoning and analysis.
- Cross-family diversity in a consensus panel.
Where it falls short
- Still pricier than the GLM-4.x line and the free flash tiers.
- Limited reproducible benchmark data; verify on your workload.
- Non-EU hosting.
Real-world use cases
- General reasoning where you want current-generation GLM quality without the flagship price.
- Document analysis and summarisation.
- Agentic tool-use pipelines.
Tokonomix benchmark snapshot
GLM-5 is newly registered on Tokonomix and not yet activated, so we have not run it through our weekly intelligence test or speed benchmark. There are no Tokonomix scores to report yet — and we will not invent any.
When it goes live, it enters the same weekly harness as every other model: identical prompts, an independent cross-family judge, and reproducible latency and cost measurements. Until then, treat the pricing and capability notes on this page as the vendor-published starting point, not as measured Tokonomix results.
EU privacy & data residency
GLM-5 is built by Zhipu AI (z.ai), a China-headquartered lab, and is served from non-EU infrastructure. This is important to state plainly: routing a prompt to this model is not an EU-data-residency or GDPR-sovereign choice, and Tokonomix will never tag it as one.
If your use case requires data to stay within the EU, pick a model whose provider is EU-hosted (for example our OVH or Azure-EU routes) rather than a GLM model. Tokonomix keeps z.ai out of every EU-only / sovereign routing set by design. Use GLM where its capability or price is the priority and cross-border processing is acceptable for that workload.
Verdict & alternatives
GLM-5 is the sensible default within the GLM-5 generation: most of the newest-generation quality at a lower price than GLM-5.2. If you need the absolute top, go GLM-5.2; if budget dominates, drop to GLM-4.6/4.5 or the free flash models.