GLM 5 Pricing - Cost Calculator

GLM-5 is Z.aiโ€™s open-weight frontier model for complex systems engineering, long-horizon reasoning, autonomous coding, agentic workflows, and technical knowledge work.

GLM 5 is a coding model from Z.ai, released 11 February 2026. It pairs a 204,800-token context window with up to 128,000 tokens of output, accepting text and returning text. Step-by-step reasoning is supported. It costs $1.00 per million input tokens through Z.ai.

Last updated Sep 29, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

๐Ÿ”’ We respect your privacy. Unsubscribe anytime.

Model Details

Released
Feb 11, 2026
Context Length
204,800
Max Output
128,000
Modalities
Text → Text
Capabilities
Tool use Structured outputs Prompt caching Open weights

Where is GLM 5 a perfect fit?

GLM-5 is Z.aiโ€™s frontier open-weight reasoning model, built for complex systems engineering and long-horizon agents, combining large-scale MoE architecture with efficient sparse attention for extended tasks. Below are the places where GLM-5 can be a perfect fit:
- Complex software engineering
- Long-running autonomous coding agents
- Large codebase analysis and refactoring
- Multi-step technical investigations
- Agentic tool-use workflows
- Systems engineering
- Long-horizon planning
- Technical research and knowledge work
- Document and spreadsheet generation
- Self-hosted frontier-model deployments
Z.ai specifically designed GLM-5 to move beyond short coding interactions toward agentic engineering, where the model can plan, execute, evaluate, and iterate on complex software tasks. It supports generating finished .docx, .pdf, and .xlsx deliverables as well.

Quick Model Estimate

(USD 1.0000 per 1M tokens)
(USD 3.2000 per 1M tokens)

Your GLM 5 Cost Estimate

๐Ÿ’ฐ Total Cost

โ€”

for 1000 input + 1000 output tokens

๐Ÿ“ฅ Input (1000 ร— $1.000000) โ€”
๐Ÿ“ค Output (1000 ร— $3.200000) โ€”

Cost Breakdown

๐Ÿ“ฅ Input ๐Ÿ“ค Output

Prices are watched for changes and checked against the provider's own page.

Pricing

Provider โ†•
Modality โ†•
Service Tier โ†•
Input Price
(per 1M tokens)
โ†•
Output Price
(per 1M tokens)
โ†•
Cached Input
(per 1M tokens)
โ†•
Context Size โ†•
View
Z.ai LogoZ.aiTextStandard$1.0000$3.2000$0.2000204,800 tokensโ†’

Benchmarks

Scores from standardized evaluations by Artificial Analysis and Design Arena. Higher is better โ€” the indices summarize overall ability, while the detailed scores break down performance on individual benchmarks.

Artificial Analysis

Higher is better ยท benchmarked by Artificial Analysis
Detailed scores
โ€”
Intelligence Index
Overall intelligence across reasoning, knowledge & math evals
โ€”
Coding Index
Coding ability across software-engineering evals
โ€”
Agentic Index
Tool use & multi-step agent task performance
Reasoning
GPQA Diamond 66.6%

Graduate-level questions in biology, chemistry & physics

Humanity's Last Exam 7.6%

Expert-level questions across many academic domains

IFBench 55.2%

Precise following of detailed instructions

ฯ„ยฒ-Bench Telecom 97.4%

Tool-using agent tasks in a telecom support setting

CritPt 0%

Unpublished physics research reasoning problems

AA-LCR 43.7%

Long-context reasoning across large inputs

Coding
Terminal-Bench Hard 39.4%

Complex command-line and terminal workflows

Knowledge
Omniscience Accuracy 23.1%

Breadth of factual knowledge across domains

Omniscience Non-Hallucination 54.7%

How reliably the model avoids fabricated answers

All variants
Variant GPQA Diamond Humanity's Last Exam IFBench ฯ„ยฒ-Bench Telecom AA-LCR Terminal-Bench Hard CritPt Omniscience Accuracy Omniscience Non-Hallucination
GLM-5 (Non-reasoning) 66.6% 7.6% 55.2% 97.4% 43.7% 39.4% 0% 23.1% 54.7%
GLM-5 (Reasoning) 82% 29.3% 72.3% 98.2% 75.7% 43.2% 2% 26.3% 64.7%

Design Arena

Elo rating by arena & category
Variant Arena Category Elo Win % Percentile Avg time (ms)
glm-5 agents androidnative 1,161 55.1% 35 โ€”
glm-5 agents fullstack 1,122 51.7% 36 โ€”
glm-5 agents godotgamedev 1,144 46.7% 42 โ€”
glm-5 agents htmlslides 1,147 45.1% 22 365,805
glm-5 agents mobileapps 1,163 51.6% 28 โ€”
glm-5 models 3d 1,247 56.3% 74 134,699
glm-5 models asciiart 1,160 47.9% 46 83,337
glm-5 models codecategories 1,255 55.5% 76 197,133
glm-5 models dataviz 1,239 53% 72 150,246
glm-5 models gamedev 1,248 57.4% 76 189,719
glm-5 models svg 1,177 54.4% 64 57,017
glm-5 models uicomponent 1,238 53.6% 70 161,437
glm-5 models website 1,255 55% 75 232,631

Compare GLM 5 with

Two models side by side: worked costs, every rate, and every benchmark figure both of them report.

Every comparison we publish →

FAQs about GLM 5

How much does GLM 5 cost per 1M tokens?

GLM 5 costs $1.00 per million input tokens, and $3.20 per million output tokens.

What does a typical workload cost with GLM 5?

1,000 requests of 2,000 input and 500 output tokens each โ€” 2,000,000 input and 500,000 output tokens in total โ€” costs $3.60 with GLM 5 at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use GLM 5?

GLM 5 is available through Z.ai, from $1.00 per million input tokens.

What is GLM 5's context window?

GLM 5 has a context window of 204,800 tokens, and returns up to 128,000 tokens in a single response.

What input types does GLM 5 support?

GLM 5 accepts text input, and returns text.