zai/glm-5
GLM 5
GLM 5 is a frontier-class, general-purpose large language model optimized for complex systems engineering and long-horizon agentic tasks. It builds on the GLM 4.5 agent-centric lineage and is designed to support multi-step reasoning, math (including AIME-style benchmarks), advanced coding, and tool-augmented workflows, with long context support suitable for sophisticated agents and enterprise applications. Typical uses include autonomous agents for software engineering, data and systems troubleshooting, operations copilots, and high-end chat assistants that must break down complex tasks, call tools reliably, and reason over long sequences of instructions or documents.
Price comparison
USD per one million tokens, using current provider data.
Provider latency
Last-hour p50 in milliseconds; lower is better.
Provider uptime
Reported availability over the last 24 hours; not a contractual SLA.
Inference providers
Compare current pricing, performance, and privacy across every available provider.
| Provider | Maximum context | Max output | p50 latency | p50 throughput | 24h uptime | Input / 1M tokens | Output / 1M tokens | Privacy |
|---|---|---|---|---|---|---|---|---|
Bedrock | 202.8K | 131.1K | 386 ms | 113.0 t/s | 99.954% | $1 | $3.2 | |
Novita | 202.8K | 131.1K | 980 ms | 129.0 t/s | 100.000% | $1 | $3.2 | |
Zai | 202.8K | 131.1K | 1,091 ms | 205.0 t/s | 100.000% | $1 | $3.2 |
What can this model do?
Available modalities and capabilities across inference providers.
Technical profile
- Model type
- Language model
- Release date
- Feb 12, 2026
- Knowledge cutoff
- —
- Regions
- —
- API specifications
- V2, V3, V4
- Temperature control
- Supported
Model identifier
zai/glm-5Related models
More options from the same creator or model category.
