GLM 4.5 AirX Pricing - Cost Calculator

GLM-4.5-AirX is Z.ai’s high-speed serving variant of GLM-4.5-Air, optimized for low-latency inference, coding, reasoning, agents, and high-concurrency applications.

GLM 4.5 AirX is a chat model from Z.ai, released 26 July 2025. It pairs a 131,072-token context window with up to 98,304 tokens of output, accepting text and returning text. Training data ends 31 December 2024, and the model reasons step by step. It costs $1.10 per million input tokens through Z.ai.

Last updated Sep 30, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

πŸ”’ We respect your privacy. Unsubscribe anytime.

Model Details

Released
Jul 26, 2025
Knowledge Cutoff
Dec 31, 2024
Context Length
131,072
Max Output
98,304
Modalities
Text → Text
Capabilities
Tool use Prompt caching Open weights

Where is GLM 4.5 AirX a perfect fit?

GLM-4.5-AirX retains the lightweight Air model’s hybrid reasoning capabilities while using faster serving infrastructure, targeting interactive applications where response latency and concurrency matter more than inference cost. The model can be a perfect fit for below use-cases:
- Low-latency assistants β€” interactive applications where response speed matters
- Agentic applications β€” fast intermediate steps in multi-step agents
- Coding assistants β€” interactive code generation and debugging
- Tool calling β€” high-frequency API/function calls
- High-concurrency workloads β€” applications serving many simultaneous requests
- Real-time chat β€” conversational applications requiring rapid responses
- Sub-agent workloads β€” fast worker model inside larger agent architectures

Quick Model Estimate

(USD 1.1000 per 1M tokens)
(USD 4.5000 per 1M tokens)

Your GLM 4.5 AirX Cost Estimate

πŸ’° Total Cost

β€”

for 1000 input + 1000 output tokens

πŸ“₯ Input (1000 Γ— $1.100000) β€”
πŸ“€ Output (1000 Γ— $4.500000) β€”

Cost Breakdown

πŸ“₯ Input πŸ“€ Output

Prices are watched for changes and checked against the provider's own page.

Pricing

Provider ↕
Modality ↕
Service Tier ↕
Input Price
(per 1M tokens)
↕
Output Price
(per 1M tokens)
↕
Cached Input
(per 1M tokens)
↕
Context Size ↕
View
Z.ai LogoZ.aiTextStandard$1.1000$4.5000$0.2200131,072 tokens→

FAQs about GLM 4.5 AirX

How much does GLM 4.5 AirX cost per 1M tokens?

GLM 4.5 AirX costs $1.10 per million input tokens, and $4.50 per million output tokens.

What does a typical workload cost with GLM 4.5 AirX?

1,000 requests of 2,000 input and 500 output tokens each β€” 2,000,000 input and 500,000 output tokens in total β€” costs $4.45 with GLM 4.5 AirX at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use GLM 4.5 AirX?

GLM 4.5 AirX is available through Z.ai, from $1.10 per million input tokens.

What is GLM 4.5 AirX's context window?

GLM 4.5 AirX has a context window of 131,072 tokens, and returns up to 98,304 tokens in a single response.

What input types does GLM 4.5 AirX support?

GLM 4.5 AirX accepts text input, and returns text.

What is GLM 4.5 AirX's knowledge cutoff?

GLM 4.5 AirX's training data runs to 31 December 2024. It has no built-in knowledge of events after that date.