GLM 4.7 Pricing - Cost Calculator

GLM-4.7 is Z.aiโ€™s advanced open-weight coding and reasoning model for agentic software development, tool use, and complex technical workflows.

GLM 4.7 is a coding model from Z.ai, released 22 December 2025. It pairs a 204,800-token context window with up to 131,072 tokens of output, taking text input and returning text. Step-by-step reasoning is supported. It costs $0.60 per million input tokens through Z.ai.

Last updated Sep 29, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

๐Ÿ”’ We respect your privacy. Unsubscribe anytime.

Model Details

Released
Dec 22, 2025
Context Length
204,800
Max Output
131,072
Modalities
Text → Text
Capabilities
Tool use Structured outputs Prompt caching Open weights

Where is GLM 4.7 a perfect fit?

GLM-4.7 delivers stronger coding, reasoning, tool use, multilingual agentic performance, and frontend generation, with interleaved and preserved thinking designed for complex, long-horizon development tasks. It can be a perfect fit for below use cases:
- Autonomous repository-level software development
- Terminal-based development โ€” debugging, implementation, and multi-step CLI workflows
- Complex reasoning โ€” mathematics, technical analysis, and difficult problem solving
- Tool-using agents โ€” web browsing, APIs, and multi-step tool orchestration
- Frontend/vibe coding โ€” modern webpages, UI components, and interactive experiences
- Multilingual development โ€” coding and agentic workflows across multiple languages
- Document and slide generation โ€” structured content with improved layout and visual quality
GLM-4.7 is primarily an agentic coding model: its major improvements over GLM-4.6 were aimed at software engineering, terminal use, tool calling, reasoning, and frontend.

Quick Model Estimate

(USD 0.6000 per 1M tokens)
(USD 2.2000 per 1M tokens)

Your GLM 4.7 Cost Estimate

๐Ÿ’ฐ Total Cost

โ€”

for 1000 input + 1000 output tokens

๐Ÿ“ฅ Input (1000 ร— $0.600000) โ€”
๐Ÿ“ค Output (1000 ร— $2.200000) โ€”

Cost Breakdown

๐Ÿ“ฅ Input ๐Ÿ“ค Output

Prices are watched for changes and checked against the provider's own page.

Pricing

Provider โ†•
Modality โ†•
Service Tier โ†•
Input Price
(per 1M tokens)
โ†•
Output Price
(per 1M tokens)
โ†•
Cached Input
(per 1M tokens)
โ†•
Context Size โ†•
View
Z.ai LogoZ.aiTextStandard$0.6000$2.2000$0.1100204,800 tokensโ†’

Benchmarks

Scores from standardized evaluations by Artificial Analysis and Design Arena. Higher is better โ€” the indices summarize overall ability, while the detailed scores break down performance on individual benchmarks.

Artificial Analysis

Higher is better ยท benchmarked by Artificial Analysis
Detailed scores
โ€”
Intelligence Index
Overall intelligence across reasoning, knowledge & math evals
โ€”
Coding Index
Coding ability across software-engineering evals
โ€”
Agentic Index
Tool use & multi-step agent task performance
Reasoning
GPQA Diamond 66.4%

Graduate-level questions in biology, chemistry & physics

Humanity's Last Exam 6.4%

Expert-level questions across many academic domains

IFBench 54.6%

Precise following of detailed instructions

ฯ„ยฒ-Bench Telecom 94.2%

Tool-using agent tasks in a telecom support setting

CritPt 0%

Unpublished physics research reasoning problems

AA-LCR 40.7%

Long-context reasoning across large inputs

Coding
Terminal-Bench Hard 30.3%

Complex command-line and terminal workflows

Knowledge
Omniscience Accuracy 23.6%

Breadth of factual knowledge across domains

Omniscience Non-Hallucination 7.1%

How reliably the model avoids fabricated answers

All variants
Variant GPQA Diamond Humanity's Last Exam IFBench ฯ„ยฒ-Bench Telecom AA-LCR Terminal-Bench Hard CritPt GDPval Omniscience Accuracy Omniscience Non-Hallucination
GLM-4.7 (Non-reasoning) 66.4% 6.4% 54.6% 94.2% 40.7% 30.3% 0% โ€” 23.6% 7.1%
GLM-4.7 (Reasoning) 85.9% 27.4% 67.9% 95.9% 71% 31.8% 1.7% 25% 29.3% 7%

Design Arena

Elo rating by arena & category
Variant Arena Category Elo Win % Percentile Avg time (ms)
glm-4.7 agents androidnative 1,128 56.2% 27 โ€”
glm-4.7 agents fullstack 1,050 45.1% 18 โ€”
glm-4.7 agents godotgamedev 1,058 35.8% 8 โ€”
glm-4.7 agents mobileapps 1,133 48.9% 24 โ€”
glm-4.7 models 3d 1,209 54.3% 64 143,428
glm-4.7 models asciiart 1,177 48.1% 56 156,554
glm-4.7 models codecategories 1,226 54.8% 67 165,180
glm-4.7 models dataviz 1,206 51.2% 60 153,971
glm-4.7 models gamedev 1,202 55.2% 63 167,669
glm-4.7 models svg 1,151 54.3% 55 90,374
glm-4.7 models uicomponent 1,207 51.1% 61 152,067
glm-4.7 models website 1,234 55.3% 69 171,664

Compare GLM 4.7 with

Two models side by side: worked costs, every rate, and every benchmark figure both of them report.

Every comparison we publish →

Other Models in the GLM 4.7 Family

FAQs about GLM 4.7

How much does GLM 4.7 cost per 1M tokens?

GLM 4.7 costs $0.60 per million input tokens, and $2.20 per million output tokens.

What does a typical workload cost with GLM 4.7?

1,000 requests of 2,000 input and 500 output tokens each โ€” 2,000,000 input and 500,000 output tokens in total โ€” costs $2.30 with GLM 4.7 at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use GLM 4.7?

GLM 4.7 is available through Z.ai, from $0.60 per million input tokens.

What is GLM 4.7's context window?

GLM 4.7 has a context window of 204,800 tokens, and returns up to 131,072 tokens in a single response.

What input types does GLM 4.7 support?

GLM 4.7 accepts text input, and returns text.