GLM-5.3 Pricing - Cost Calculator

GLM-5.3 is Z.aiโ€™s frontier open-weight model for advanced coding, autonomous agents, cybersecurity, long-horizon reasoning, and complex software engineering workflows.

Released 14 August 2026, GLM-5.3 is a text model from Z.ai. It pairs a 1,000,000-token context window with up to 128,000 tokens of output, taking text input and returning text. Step-by-step reasoning is supported. It costs $1.40 per million input tokens through Perplexity.

Last updated Sep 15, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

๐Ÿ”’ We respect your privacy. Unsubscribe anytime.

Model Details

Released
Aug 14, 2026
Context Length
1,000,000
Max Output
128,000
Modalities
Text Text
Capabilities
Tool use Structured outputs Prompt caching Code execution Open weights

Where is GLM-5.3 a perfect fit?

GLM-5.3 is Z.aiโ€™s frontier open-weight model focused heavily on coding and autonomous agents, delivering major post-training gains for long-horizon software engineering, cybersecurity, and complex technical workflows.
- Autonomous software engineering agents
- Repository-scale coding and codebase maintenance
- Long-running terminal/CLI agents
- Cybersecurity vulnerability research and defensive analysis
- Complex multi-step technical tasks
- Agentic coding with tool use
- Long-context code and documentation analysis
- Self-hosted frontier-model deployments

GLM-5.3 is essentially a heavily post-trained GLM-5.2, optimized to turn a strong open-weight foundation into a much more capable coding and autonomous-agent model. Its standout combination is 1M context + open weights + frontier-level coding/agentic performance. Z.ai describes GLM-5.3 as its most capable open-weights model for coding, reporting a roughly 50% improvement over GLM-5.2 on its internal Code Bench.

Quick Model Estimate

(USD 1.4000 per 1M tokens)
(USD 4.4000 per 1M tokens)

Your GLM-5.3 Cost Estimate

๐Ÿ’ฐ Total Cost

โ€”

for 1000 input + 1000 output tokens

๐Ÿ“ฅ Input (1000 ร— $1.400000) โ€”
๐Ÿ“ค Output (1000 ร— $4.400000) โ€”

Cost Breakdown

๐Ÿ“ฅ Input ๐Ÿ“ค Output

Prices updated daily from official provider data.

Pricing

Provider โ†•
Modality โ†•
Service Tier โ†•
Input Price
(per 1M tokens)
โ†•
Output Price
(per 1M tokens)
โ†•
Cached Input
(per 1M tokens)
โ†•
Context Size โ†•
View
Perplexity LogoPerplexityTextStandard$1.4000$4.4000$0.26001,000,000 tokensโ†’

Benchmarks

Scores from standardized evaluations by Artificial Analysis and Design Arena. Higher is better โ€” the indices summarize overall ability, while the detailed scores break down performance on individual benchmarks.

Artificial Analysis

Higher is better ยท benchmarked by Artificial Analysis
Detailed scores
GLM-5.3 (max)
44.9
Intelligence Index
Overall intelligence across reasoning, knowledge & math evals
Better than 63% of 9 models
74.8
Coding Index
Coding ability across software-engineering evals
Better than 56% of 10 models
53.4
Agentic Index
Tool use & multi-step agent task performance
Better than 88% of 9 models
Reasoning
GPQA Diamond 91.7%

Graduate-level questions in biology, chemistry & physics

Humanity's Last Exam 42.3%

Expert-level questions across many academic domains

CritPt 19.1%

Unpublished physics research reasoning problems

AA-LCR 79.7%

Long-context reasoning across large inputs

Coding
SciCode 59%

Research-level scientific coding tasks

Knowledge
GDPval 57.8%

Economically valuable, real-world knowledge work

Omniscience Accuracy 33.9%

Breadth of factual knowledge across domains

Omniscience Non-Hallucination 70.4%

How reliably the model avoids fabricated answers

Design Arena

Elo rating by arena & category
Variant Arena Category Elo Win % Percentile Avg time (ms)
glm-5.3 agents htmlslides 1,189 39.1% 61 โ€”
glm-5.3 agents mobileapps 1,229 54.4% 69 โ€”
glm-5.3 agents python-pptxslides 1,259 51.6% 81 โ€”
cortado models 3d 1,393 67.8% 97 537,132
cortado models codecategories 1,330 57.7% 96 407,296
cortado models dataviz 1,262 51% 80 375,944
cortado models gamedev 1,376 64.1% 98 580,150
cortado models svg 1,325 59.8% 95 244,308
cortado models uicomponent 1,344 59.5% 96 359,679
cortado models website 1,317 55.8% 95 369,976

FAQs about GLM-5.3

How much does GLM-5.3 cost per 1M tokens?

GLM-5.3 costs $1.40 per million input tokens, and $4.40 per million output tokens.

What does a typical workload cost with GLM-5.3?

1,000 requests of 2,000 input and 500 output tokens each โ€” 2,000,000 input and 500,000 output tokens in total โ€” costs $5.00 with GLM-5.3 at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use GLM-5.3?

GLM-5.3 is available through Perplexity, from $1.40 per million input tokens.

What is GLM-5.3's context window?

GLM-5.3 has a context window of 1,000,000 tokens, and returns up to 128,000 tokens in a single response.

What input types does GLM-5.3 support?

GLM-5.3 accepts text input, and returns text.