Cheaper overall
GLM 4.7 FlashX
24% lower blended rate
Higher benchmark scores
GPT-6 Luna
leads on 5 of 5 figures
Bigger context window
GPT-6 Luna
1,050,000 vs 202,752 tokens
Where each one wins
GPT-6 Luna
- Larger context window 1,050,000 tokens
- Higher expert-exam performance (Humanity's Last Exam) 38.5 vs 7.6
- Higher long-context reasoning (AA-LCR) 83.3 vs 41.7
- Higher physics research reasoning (CritPt) 19.4 vs 0.3
- Ahead on 2 more benchmark figures
- Offers Batch, Flex and Priority pricing not listed for GLM 4.7 FlashX
GLM 4.7 FlashX
- Cheaper input tokens $0.07 vs $0.10, 30% less
- Cheaper output tokens $0.40 vs $0.50, 20% less
- Cheaper on all four workloads
What four workloads cost
One run, and the same run a thousand times.
GLM 4.7 FlashX is cheaper on all four workloads, by much the same margin each time โ about 1.3 times, from $0.0003 against $0.0004 on a chat turn to $0.0144 against $0.0170 on a long document.
| Workload | GPT-6 Luna | GLM 4.7 FlashX | ร1,000 | Difference |
|---|---|---|---|---|
| Chat turn 1,000 in / 500 out | $0.0004 | $0.0003 | $0.3500 / $0.2700 | GLM 4.7 FlashX is 23% cheaper |
| RAG answer 10,000 in / 800 out | $0.0014 | $0.0010 | $1.40 / $1.02 | GLM 4.7 FlashX is 27% cheaper |
| Agent loop 32,000 in / 700 out ร 20 calls | $0.0170 | $0.0144 | $17.00 / $14.40 | GLM 4.7 FlashX is 15% cheaper |
| Long document 400,000 in / 3,000 out | GLM 4.7 FlashX's context window holds 202,752 tokens, so a 400,000-token prompt does not fit. | |||
Price per 1M tokens
Above 272,000 input tokens the rates change, and the new rate applies to the whole request, not just the tokens past that point: GPT-6 Luna's input goes from $0.1000 to $0.2000. It does not change which of the two is cheaper.
| Context band | Price | GPT-6 Luna | GLM 4.7 FlashX | Difference |
|---|---|---|---|---|
| Prompts up to 272,000 tokens | Input | $0.10 | $0.07 | GLM 4.7 FlashX is 30% cheaper |
| Prompts up to 272,000 tokens | Cached input | $0.01 | $0.01 | Same |
| Prompts up to 272,000 tokens | Output | $0.50 | $0.40 | GLM 4.7 FlashX is 20% cheaper |
| Prompts over 272,000 tokens | Input | $0.20 | $0.07 | GLM 4.7 FlashX is 65% cheaper |
| Prompts over 272,000 tokens | Cached input | $0.02 | $0.01 | GLM 4.7 FlashX is 50% cheaper |
| Prompts over 272,000 tokens | Output | $0.75 | $0.40 | GLM 4.7 FlashX is 47% cheaper |
Blended: GPT-6 Luna $0.20, GLM 4.7 FlashX $0.1525 per 1M โ one rate at 3:1 input to output, for comparing two models at a glance.
GPT-6 Luna priced by OpenAI, GLM 4.7 FlashX by Z.ai. Standard tier, pay-as-you-go.
Other pricing tiers
Only GPT-6 Luna lists Batch, Flex and Priority, at up to 50% off its own standard rate; GLM 4.7 FlashX offers none of them.
Every tier either model sells is here, so this is the whole price sheet. A saving is measured against that model's own standard rate โ how we read tiers.
| Tier | GPT-6 Luna | GLM 4.7 FlashX | Cheaper |
|---|---|---|---|
| Batch only on GPT-6 Luna | $0.05 in / $0.25 out 50% off Standard | Not offered | — |
| Flex only on GPT-6 Luna | $0.05 in / $0.25 out 50% off Standard | Not offered | — |
| Priority only on GPT-6 Luna | $0.20 in / $1.00 out | Not offered | — |
Benchmarks
GPT-6 Luna leads on all 5 figures, furthest ahead on long-context reasoning (AA-LCR), by 41.6 points.
Expert-exam performance
Humanity's Last Exam โ Expert-level questions across many academic domains
GPT-6 Luna ahead by 30.9
Long-context reasoning
AA-LCR โ Long-context reasoning across large inputs
GPT-6 Luna ahead by 41.6
Physics research reasoning
CritPt โ Unpublished physics research reasoning problems
GPT-6 Luna ahead by 19.1
Factual accuracy
Omniscience Accuracy โ Breadth of factual knowledge across domains
GPT-6 Luna ahead by 27.6
Answer reliability
Omniscience Non-Hallucination โ How reliably the model avoids fabricated answers
GPT-6 Luna ahead by 17.2
Each figure is that model's best published run, with the effort level named beside it โ where these scores come from.
Specifications
GPT-6 Luna holds 847,248 more input tokens in one request. GPT-6 Luna also takes file and image. Only GPT-6 Luna lists web search, code execution and computer use; only GLM 4.7 FlashX lists open weights.
| Specification | GPT-6 Luna | GLM 4.7 FlashX |
|---|---|---|
| Context window | 1,050,000 tokens | 202,752 tokens |
| Max output | 128,000 tokens | 128,000 tokens |
| Takes in | Text, image, file | Text |
| Puts out | Text | Text |
| Knowledge cutoff | 18 May 2026 | โ |
| Released | 22 Sep 2026 | 19 Jan 2026 |
| Status | Active | Active |
| Sold by | OpenAI | Z.ai |
| Tool use | Yes | Yes |
| Structured outputs | Yes | Yes |
| Web search | Yes | Not listed |
| Prompt caching | Yes | Yes |
| Code execution | Yes | Not listed |
| Computer use | Yes | Not listed |
| Open weights | Not listed | Yes |
| Reasoning | Yes | Yes |
FAQs
- Is GPT-6 Luna cheaper than GLM 4.7 FlashX?
- No โ the other way round. A 1,000-token prompt with a 500-token reply costs $0.0004 on GPT-6 Luna and $0.0003 on GLM 4.7 FlashX, and the same model is cheaper on every workload on this page.
- How much do GPT-6 Luna and GLM 4.7 FlashX cost per 1M tokens?
- GPT-6 Luna costs $0.10 for input and $0.50 for output. GLM 4.7 FlashX costs $0.07 and $0.40. Standard pay-as-you-go rates, as of 22 Sep 2026.
- Which is better, GPT-6 Luna or GLM 4.7 FlashX?
- On published benchmarks, GPT-6 Luna. It leads on 5 of the 5 figures both models report, including expert-exam performance (Humanity's Last Exam), where it scores 38.5 against 7.6.
- Does GPT-6 Luna or GLM 4.7 FlashX offer batch pricing?
- GPT-6 Luna only. Its batch input costs $0.05, 50% off its standard rate, and GLM 4.7 FlashX lists no batch tier at all.
- Does GPT-6 Luna or GLM 4.7 FlashX have a bigger context window?
- GPT-6 Luna, at 1,050,000 tokens against 202,752. That is 847,248 more input tokens in a single request.