Model Details
Where is GLM 4 32B-0414-128K a perfect fit?
- Coding — code generation, debugging, and software development
- Reasoning — mathematical and logical problem solving
- General dialogue — Chinese and English conversational applications
- Long-context analysis — large documents and technical material
- AI assistants — general-purpose local assistants
- Self-hosted AI — organizations requiring local model deployment
- Research and experimentation — fine-tuning and model-development work
Quick Model Estimate
Your GLM 4 32B-0414-128K Cost Estimate
💰 Total Cost
—
for 1000 input + 1000 output tokens
Cost Breakdown
Prices are watched for changes and checked against the provider's own page.
Pricing
FAQs about GLM 4 32B-0414-128K
How much does GLM 4 32B-0414-128K cost per 1M tokens?
GLM 4 32B-0414-128K costs $0.10 per million input tokens, and $0.10 per million output tokens.
What does a typical workload cost with GLM 4 32B-0414-128K?
1,000 requests of 2,000 input and 500 output tokens each — 2,000,000 input and 500,000 output tokens in total — costs $0.25 with GLM 4 32B-0414-128K at its lowest rates. Use the calculator on this page for your own volumes.
Where can I use GLM 4 32B-0414-128K?
GLM 4 32B-0414-128K is available through Z.ai, from $0.10 per million input tokens.
What is GLM 4 32B-0414-128K's context window?
GLM 4 32B-0414-128K has a context window of 128,000 tokens.
What input types does GLM 4 32B-0414-128K support?
GLM 4 32B-0414-128K accepts text input, and returns text.
What is GLM 4 32B-0414-128K's knowledge cutoff?
GLM 4 32B-0414-128K's training data runs to 30 June 2024. It has no built-in knowledge of events after that date.
