Cheaper overall
GPT-5.6 Luna
71% lower blended rate
Higher benchmark scores
GPT-5.6 Luna
leads on 5 of 6 figures
Bigger context window
GPT-5.6 Luna
1,050,000 vs 204,800 tokens
Cheaper cached input
GPT-5.6 Luna
$0.02 vs $0.20, 90% less
Where each one wins
GPT-5.6 Luna
- Cheaper input tokens $0.20 vs $1.00, 80% less
- Cheaper output tokens $1.20 vs $3.20, 63% less
- Cheaper cached input $0.02 vs $0.20, 90% less
- Cheaper on all four workloads
- Larger context window 1,050,000 tokens
- Higher graduate-level science (GPQA Diamond) 91.1 vs 82
- Higher expert-exam performance (Humanity's Last Exam) 39.5 vs 29.3
- Higher long-context reasoning (AA-LCR) 83.7 vs 75.7
- Ahead on 2 more benchmark figures
- Offers Batch, Flex and Priority pricing not listed for GLM 5
GLM 5
- Higher answer reliability (Omniscience Non-Hallucination) 64.7 vs 25
What four workloads cost
One run, and the same run a thousand times.
GPT-5.6 Luna is cheaper on all four workloads, by much the same margin each time โ about 4.4 times, from $0.0008 against $0.0026 on a chat turn to $0.0368 against $0.2048 on a long document.
| Workload | GPT-5.6 Luna | GLM 5 | ร1,000 | Difference |
|---|---|---|---|---|
| Chat turn 1,000 in / 500 out | $0.0008 | $0.0026 | $0.8000 / $2.60 | GPT-5.6 Luna is 69% cheaper |
| RAG answer 10,000 in / 800 out | $0.0030 | $0.0126 | $2.96 / $12.56 | GPT-5.6 Luna is 76% cheaper |
| Agent loop 32,000 in / 700 out ร 20 calls | $0.0368 | $0.2048 | $36.80 / $204.80 | GPT-5.6 Luna is 82% cheaper |
| Long document 400,000 in / 3,000 out | GLM 5's context window holds 204,800 tokens, so a 400,000-token prompt does not fit. | |||
Price per 1M tokens
Above 272,000 input tokens the rates change, and the new rate applies to the whole request, not just the tokens past that point: GPT-5.6 Luna's input goes from $0.2000 to $0.4000. It does not change which of the two is cheaper.
| Context band | Price | GPT-5.6 Luna | GLM 5 | Difference |
|---|---|---|---|---|
| Prompts up to 272,000 tokens | Input | $0.20 | $1.00 | GPT-5.6 Luna is 80% cheaper |
| Prompts up to 272,000 tokens | Cached input | $0.02 | $0.20 | GPT-5.6 Luna is 90% cheaper |
| Prompts up to 272,000 tokens | Output | $1.20 | $3.20 | GPT-5.6 Luna is 63% cheaper |
| Prompts over 272,000 tokens | Input | $0.40 | $1.00 | GPT-5.6 Luna is 60% cheaper |
| Prompts over 272,000 tokens | Cached input | $0.04 | $0.20 | GPT-5.6 Luna is 80% cheaper |
| Prompts over 272,000 tokens | Output | $1.80 | $3.20 | GPT-5.6 Luna is 44% cheaper |
Blended: GPT-5.6 Luna $0.45, GLM 5 $1.55 per 1M โ one rate at 3:1 input to output, for comparing two models at a glance.
GPT-5.6 Luna priced by OpenAI, GLM 5 by Z.ai. Standard tier, pay-as-you-go.
Other pricing tiers
Only GPT-5.6 Luna lists Batch, Flex and Priority, at up to 50% off its own standard rate; GLM 5 offers none of them.
Every tier either model sells is here, so this is the whole price sheet. A saving is measured against that model's own standard rate โ how we read tiers.
| Tier | GPT-5.6 Luna | GLM 5 | Cheaper |
|---|---|---|---|
| Batch only on GPT-5.6 Luna | $0.10 in / $0.60 out 50% off Standard | Not offered | — |
| Flex only on GPT-5.6 Luna | $0.10 in / $0.60 out 50% off Standard | Not offered | — |
| Priority only on GPT-5.6 Luna | $0.40 in / $2.40 out | Not offered | — |
Benchmarks
They share 6 figures: GPT-5.6 Luna leads on 5, GLM 5 on 1. The widest gap is 39.7 points, on answer reliability (Omniscience Non-Hallucination).
Graduate-level science
GPQA Diamond โ Graduate-level questions in biology, chemistry & physics
GPT-5.6 Luna ahead by 9.1
Expert-exam performance
Humanity's Last Exam โ Expert-level questions across many academic domains
GPT-5.6 Luna ahead by 10.2
Long-context reasoning
AA-LCR โ Long-context reasoning across large inputs
GPT-5.6 Luna ahead by 8
Physics research reasoning
CritPt โ Unpublished physics research reasoning problems
GPT-5.6 Luna ahead by 18.6
Factual accuracy
Omniscience Accuracy โ Breadth of factual knowledge across domains
GPT-5.6 Luna ahead by 16.4
Answer reliability
Omniscience Non-Hallucination โ How reliably the model avoids fabricated answers
GLM 5 ahead by 39.7
Each figure is that model's best published run, with the effort level named beside it โ where these scores come from.
Specifications
GPT-5.6 Luna holds 845,200 more input tokens in one request. GPT-5.6 Luna also takes file and image. Only GPT-5.6 Luna lists code execution; only GLM 5 lists open weights.
| Specification | GPT-5.6 Luna | GLM 5 |
|---|---|---|
| Context window | 1,050,000 tokens | 204,800 tokens |
| Max output | 128,000 tokens | 128,000 tokens |
| Takes in | Text, image, file | Text |
| Puts out | Text | Text |
| Knowledge cutoff | 16 Feb 2026 | โ |
| Released | 9 Jul 2026 | 11 Feb 2026 |
| Status | Active | Active |
| Sold by | OpenAI, Perplexity | Z.ai |
| Tool use | Yes | Yes |
| Structured outputs | Yes | Yes |
| Web search | Not listed | Not listed |
| Prompt caching | Yes | Yes |
| Code execution | Yes | Not listed |
| Computer use | Not listed | Not listed |
| Open weights | Not listed | Yes |
| Reasoning | Yes | Yes |
FAQs
- Is GPT-5.6 Luna cheaper than GLM 5?
- Yes. A 1,000-token prompt with a 500-token reply costs $0.0008 on GPT-5.6 Luna and $0.0026 on GLM 5, and the same model is cheaper on every workload on this page.
- How much do GPT-5.6 Luna and GLM 5 cost per 1M tokens?
- GPT-5.6 Luna costs $0.20 for input and $1.20 for output. GLM 5 costs $1.00 and $3.20. Standard pay-as-you-go rates, as of 30 Jul 2026.
- Which is better, GPT-5.6 Luna or GLM 5?
- On published benchmarks, GPT-5.6 Luna. It leads on 5 of the 6 figures both models report, including graduate-level science (GPQA Diamond), where it scores 91.1 against 82.
- Does GPT-5.6 Luna or GLM 5 offer batch pricing?
- GPT-5.6 Luna only. Its batch input costs $0.10, 50% off its standard rate, and GLM 5 lists no batch tier at all.
- Does GPT-5.6 Luna or GLM 5 have a bigger context window?
- GPT-5.6 Luna, at 1,050,000 tokens against 204,800. That is 845,200 more input tokens in a single request.