Model Details
Where is GLM 5 a perfect fit?
- Complex software engineering
- Long-running autonomous coding agents
- Large codebase analysis and refactoring
- Multi-step technical investigations
- Agentic tool-use workflows
- Systems engineering
- Long-horizon planning
- Technical research and knowledge work
- Document and spreadsheet generation
- Self-hosted frontier-model deployments
Z.ai specifically designed GLM-5 to move beyond short coding interactions toward agentic engineering, where the model can plan, execute, evaluate, and iterate on complex software tasks. It supports generating finished .docx, .pdf, and .xlsx deliverables as well.
Quick Model Estimate
Your GLM 5 Cost Estimate
๐ฐ Total Cost
โ
for 1000 input + 1000 output tokens
Cost Breakdown
Prices are watched for changes and checked against the provider's own page.
Pricing
Benchmarks
Scores from standardized evaluations by Artificial Analysis and Design Arena. Higher is better โ the indices summarize overall ability, while the detailed scores break down performance on individual benchmarks.
Artificial Analysis
Higher is better ยท benchmarked by Artificial AnalysisGraduate-level questions in biology, chemistry & physics
Expert-level questions across many academic domains
Precise following of detailed instructions
Tool-using agent tasks in a telecom support setting
Unpublished physics research reasoning problems
Long-context reasoning across large inputs
Complex command-line and terminal workflows
Breadth of factual knowledge across domains
How reliably the model avoids fabricated answers
| Variant | GPQA Diamond | Humanity's Last Exam | IFBench | ฯยฒ-Bench Telecom | AA-LCR | Terminal-Bench Hard | CritPt | Omniscience Accuracy | Omniscience Non-Hallucination |
|---|---|---|---|---|---|---|---|---|---|
| GLM-5 (Non-reasoning) | 66.6% | 7.6% | 55.2% | 97.4% | 43.7% | 39.4% | 0% | 23.1% | 54.7% |
| GLM-5 (Reasoning) | 82% | 29.3% | 72.3% | 98.2% | 75.7% | 43.2% | 2% | 26.3% | 64.7% |
Design Arena
Elo rating by arena & category| Variant | Arena | Category | Elo | Win % | Percentile | Avg time (ms) |
|---|---|---|---|---|---|---|
| glm-5 | agents | androidnative | 1,161 | 55.1% | 35 | โ |
| glm-5 | agents | fullstack | 1,122 | 51.7% | 36 | โ |
| glm-5 | agents | godotgamedev | 1,144 | 46.7% | 42 | โ |
| glm-5 | agents | htmlslides | 1,147 | 45.1% | 22 | 365,805 |
| glm-5 | agents | mobileapps | 1,163 | 51.6% | 28 | โ |
| glm-5 | models | 3d | 1,247 | 56.3% | 74 | 134,699 |
| glm-5 | models | asciiart | 1,160 | 47.9% | 46 | 83,337 |
| glm-5 | models | codecategories | 1,255 | 55.5% | 76 | 197,133 |
| glm-5 | models | dataviz | 1,239 | 53% | 72 | 150,246 |
| glm-5 | models | gamedev | 1,248 | 57.4% | 76 | 189,719 |
| glm-5 | models | svg | 1,177 | 54.4% | 64 | 57,017 |
| glm-5 | models | uicomponent | 1,238 | 53.6% | 70 | 161,437 |
| glm-5 | models | website | 1,255 | 55% | 75 | 232,631 |
Compare GLM 5 with
Two models side by side: worked costs, every rate, and every benchmark figure both of them report.
FAQs about GLM 5
How much does GLM 5 cost per 1M tokens?
GLM 5 costs $1.00 per million input tokens, and $3.20 per million output tokens.
What does a typical workload cost with GLM 5?
1,000 requests of 2,000 input and 500 output tokens each โ 2,000,000 input and 500,000 output tokens in total โ costs $3.60 with GLM 5 at its lowest rates. Use the calculator on this page for your own volumes.
Where can I use GLM 5?
GLM 5 is available through Z.ai, from $1.00 per million input tokens.
What is GLM 5's context window?
GLM 5 has a context window of 204,800 tokens, and returns up to 128,000 tokens in a single response.
What input types does GLM 5 support?
GLM 5 accepts text input, and returns text.
