Provider Details
- Headquarter: Beijing, China
- Notable Models: GLM-5.2, GLM-5.1, GLM-5V-Turbo
Model Release Timeline
- Open, high-performance models: GLM models deliver competitive reasoning, coding, and multimodal capabilities with strong price-performance.
- Developer-first APIs: Simple integration for AI assistants, enterprise applications, coding tools, and autonomous AI agents.
- Built for scale: Supports high-throughput inference and enterprise deployments with flexible model options.
- Cost optimization: ModelCosts.com helps forecast API expenses, compare GLM variants, and receive pricing alerts before costs increase.
- Competitive advantage: Open-weight strategy, rapid model iteration, and efficient inference make Z.ai an attractive choice for organizations seeking lower AI infrastructure costs without sacrificing capability.
Quick Model Estimate
Your Cost Estimate
💰 Total Cost
—
for 1000 input + 1000 output tokens
Cost Breakdown
Prices updated daily from official provider data.
Pricing
|
Model
↕
|
Modality
↕
|
Service Tier
↕
|
Input Price
(per 1M tokens) ↕ |
Output Price
(per 1M tokens) ↕ |
Cached Input
(per 1M tokens) ↕ |
Context Window
↕
|
View
|
|---|---|---|---|---|---|---|---|
| GLM-5.3-Flash | Text | Standard | $0.1500 | $0.5000 | $0.0300 | 1,000,000 tokens | → |
| GLM-5.3 | Text | Standard | $1.4000 | $4.4000 | $0.2600 | 1,000,000 tokens | → |
| GLM-5.2 | Text | Standard | $1.4000 | $4.4000 | $0.2600 | 1,000,000 tokens | → |
| GLM 5.1 | Text | Standard | $1.4000 | $4.4000 | $0.2600 | 200,000 tokens | → |
| GLM 5.3 FlashX | Text | Standard | $0.3700 | $1.2500 | $0.0750 | 1,048,576 tokens | → |