Model Details
Where is GLM 4.7 FlashX a perfect fit?
- Low-latency applications โ fast interactive AI responses
- Coding agents โ lightweight autonomous software-development agents
- Tool-calling workflows โ high-volume function and API calls
- Reasoning tasks โ mathematics, analysis, and structured problem solving
- Production chat โ inexpensive high-throughput conversational workloads
- Sub-agents โ economical worker models inside larger agent systems
- Self-hosted AI โ efficient deployment compared with large frontier models
Quick Model Estimate
Your GLM 4.7 FlashX Cost Estimate
๐ฐ Total Cost
โ
for 1000 input + 1000 output tokens
Cost Breakdown
Prices are watched for changes and checked against the provider's own page.
Pricing
Benchmarks
Scores from standardized evaluations by Artificial Analysis and Design Arena. Higher is better โ the indices summarize overall ability, while the detailed scores break down performance on individual benchmarks.
Artificial Analysis
Higher is better ยท benchmarked by Artificial AnalysisGraduate-level questions in biology, chemistry & physics
Expert-level questions across many academic domains
Precise following of detailed instructions
Tool-using agent tasks in a telecom support setting
Unpublished physics research reasoning problems
Long-context reasoning across large inputs
Complex command-line and terminal workflows
Breadth of factual knowledge across domains
How reliably the model avoids fabricated answers
| Variant | GPQA Diamond | Humanity's Last Exam | IFBench | ฯยฒ-Bench Telecom | AA-LCR | Terminal-Bench Hard | CritPt | Omniscience Accuracy | Omniscience Non-Hallucination |
|---|---|---|---|---|---|---|---|---|---|
| GLM-4.7-Flash (Non-reasoning) | 45.2% | 5% | 46.3% | 91.8% | 20.3% | 3.8% | 0% | 13.1% | 5.7% |
| GLM-4.7-Flash (Reasoning) | 58.1% | 7.6% | 60.8% | 98.8% | 41.7% | 22% | 0.3% | 16.2% | 6.1% |
Design Arena
Elo rating by arena & category| Variant | Arena | Category | Elo | Win % | Percentile | Avg time (ms) |
|---|---|---|---|---|---|---|
| glm-4.7-flash | models | 3d | 1,139 | 51.2% | 42 | 116,926 |
| glm-4.7-flash | models | codecategories | 1,187 | 53.1% | 55 | 150,068 |
| glm-4.7-flash | models | dataviz | 1,134 | 45.3% | 30 | 126,750 |
| glm-4.7-flash | models | gamedev | 1,149 | 49.7% | 41 | 153,698 |
| glm-4.7-flash | models | svg | 1,046 | 44.2% | 16 | 81,729 |
| glm-4.7-flash | models | uicomponent | 1,216 | 57.6% | 64 | 116,336 |
| glm-4.7-flash | models | website | 1,202 | 54% | 58 | 159,006 |
Compare GLM 4.7 FlashX with
Two models side by side: worked costs, every rate, and every benchmark figure both of them report.
Other Models in the GLM 4.7 Family
FAQs about GLM 4.7 FlashX
How much does GLM 4.7 FlashX cost per 1M tokens?
GLM 4.7 FlashX costs $0.07 per million input tokens, and $0.40 per million output tokens.
What does a typical workload cost with GLM 4.7 FlashX?
1,000 requests of 2,000 input and 500 output tokens each โ 2,000,000 input and 500,000 output tokens in total โ costs $0.34 with GLM 4.7 FlashX at its lowest rates. Use the calculator on this page for your own volumes.
Where can I use GLM 4.7 FlashX?
GLM 4.7 FlashX is available through Z.ai, from $0.07 per million input tokens.
What is GLM 4.7 FlashX's context window?
GLM 4.7 FlashX has a context window of 202,752 tokens, and returns up to 128,000 tokens in a single response.
What input types does GLM 4.7 FlashX support?
GLM 4.7 FlashX accepts text input, and returns text.
