GLM 4 32B-0414-128K Pricing - Cost Calculator

GLM-4-32B-0414-128K is Z.ai’s 32-billion-parameter open-weight model for reasoning, coding, dialogue, and long-context technical applications.

GLM 4 32B-0414-128K, released 14 April 2025, is a reasoning model from Z.ai. It costs $0.10 per million input tokens through Z.ai. The context window holds 128,000 tokens; the model takes text input and returns text. It reasons step by step, and its training data ends 30 June 2024.

Last updated Oct 1, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

🔒 We respect your privacy. Unsubscribe anytime.

Model Details

Released
Apr 14, 2025
Knowledge Cutoff
Jun 30, 2024
Context Length
128,000
Modalities
Text → Text
Capabilities
Code execution Open weights

Where is GLM 4 32B-0414-128K a perfect fit?

GLM-4-32B-0414-128K extends Z.ai’s GLM-4 architecture with stronger reasoning and coding capabilities, supporting dialogue, problem solving, local deployment, and long-context workloads under MIT licensing. The model can be a perfect fit for:
- Coding — code generation, debugging, and software development
- Reasoning — mathematical and logical problem solving
- General dialogue — Chinese and English conversational applications
- Long-context analysis — large documents and technical material
- AI assistants — general-purpose local assistants
- Self-hosted AI — organizations requiring local model deployment
- Research and experimentation — fine-tuning and model-development work

Quick Model Estimate

(USD 0.1000 per 1M tokens)
(USD 0.1000 per 1M tokens)

Your GLM 4 32B-0414-128K Cost Estimate

💰 Total Cost

—

for 1000 input + 1000 output tokens

📥 Input (1000 × $0.100000) —
📤 Output (1000 × $0.100000) —

Cost Breakdown

📥 Input 📤 Output

Prices are watched for changes and checked against the provider's own page.

Pricing

Provider ↕
Modality ↕
Service Tier ↕
Input Price
(per 1M tokens)
↕
Output Price
(per 1M tokens)
↕
Context Size ↕
View
Z.ai LogoZ.aiTextStandard$0.1000$0.1000128,000 tokens→

FAQs about GLM 4 32B-0414-128K

How much does GLM 4 32B-0414-128K cost per 1M tokens?

GLM 4 32B-0414-128K costs $0.10 per million input tokens, and $0.10 per million output tokens.

What does a typical workload cost with GLM 4 32B-0414-128K?

1,000 requests of 2,000 input and 500 output tokens each — 2,000,000 input and 500,000 output tokens in total — costs $0.25 with GLM 4 32B-0414-128K at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use GLM 4 32B-0414-128K?

GLM 4 32B-0414-128K is available through Z.ai, from $0.10 per million input tokens.

What is GLM 4 32B-0414-128K's context window?

GLM 4 32B-0414-128K has a context window of 128,000 tokens.

What input types does GLM 4 32B-0414-128K support?

GLM 4 32B-0414-128K accepts text input, and returns text.

What is GLM 4 32B-0414-128K's knowledge cutoff?

GLM 4 32B-0414-128K's training data runs to 30 June 2024. It has no built-in knowledge of events after that date.