GPT-6 Luna vs GLM 4.7 FlashX

GLM 4.7 FlashX is the cheaper of the two, though not evenly: its input costs 30% less than GPT-6 Luna's, $0.07 per 1M against $0.10, while its output is only 20% cheaper, $0.40 against $0.50. A single chat turn costs a fraction of a cent on both, so this only becomes real money at volume: a thousand agent loops costs $14.40 against $17.00.

What the extra 31% buys is 30.9 points of expert-exam performance (Humanity's Last Exam): 38.5 (max) against 7.6 (GLM-4.7-Flash). GLM 4.7 FlashX is the one to pick unless you need GPT-6 Luna's lead on expert-exam performance.

Standard pay-as-you-go prices, as of 22 Sep 2026. Last updated 29 Sep 2026.

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

๐Ÿ”’ We respect your privacy. Unsubscribe anytime.

Cheaper overall

GLM 4.7 FlashX

24% lower blended rate

Higher benchmark scores

GPT-6 Luna

leads on 5 of 5 figures

Bigger context window

GPT-6 Luna

1,050,000 vs 202,752 tokens

Where each one wins

GPT-6 Luna

  • Larger context window 1,050,000 tokens
  • Higher expert-exam performance (Humanity's Last Exam) 38.5 vs 7.6
  • Higher long-context reasoning (AA-LCR) 83.3 vs 41.7
  • Higher physics research reasoning (CritPt) 19.4 vs 0.3
  • Ahead on 2 more benchmark figures
  • Offers Batch, Flex and Priority pricing not listed for GLM 4.7 FlashX

GLM 4.7 FlashX

  • Cheaper input tokens $0.07 vs $0.10, 30% less
  • Cheaper output tokens $0.40 vs $0.50, 20% less
  • Cheaper on all four workloads

What four workloads cost

One run, and the same run a thousand times.

GLM 4.7 FlashX is cheaper on all four workloads, by much the same margin each time โ€” about 1.3 times, from $0.0003 against $0.0004 on a chat turn to $0.0144 against $0.0170 on a long document.

Workload GPT-6 Luna GLM 4.7 FlashX ร—1,000 Difference
Chat turn 1,000 in / 500 out $0.0004 $0.0003 $0.3500 / $0.2700 GLM 4.7 FlashX is 23% cheaper
RAG answer 10,000 in / 800 out $0.0014 $0.0010 $1.40 / $1.02 GLM 4.7 FlashX is 27% cheaper
Agent loop 32,000 in / 700 out ร— 20 calls $0.0170 $0.0144 $17.00 / $14.40 GLM 4.7 FlashX is 15% cheaper
Long document 400,000 in / 3,000 out GLM 4.7 FlashX's context window holds 202,752 tokens, so a 400,000-token prompt does not fit.

Price your own token counts for these two โ†’

Price per 1M tokens

Above 272,000 input tokens the rates change, and the new rate applies to the whole request, not just the tokens past that point: GPT-6 Luna's input goes from $0.1000 to $0.2000. It does not change which of the two is cheaper.

Context band Price GPT-6 Luna GLM 4.7 FlashX Difference
Prompts up to 272,000 tokens Input $0.10 $0.07 GLM 4.7 FlashX is 30% cheaper
Prompts up to 272,000 tokens Cached input $0.01 $0.01 Same
Prompts up to 272,000 tokens Output $0.50 $0.40 GLM 4.7 FlashX is 20% cheaper
Prompts over 272,000 tokens Input $0.20 $0.07 GLM 4.7 FlashX is 65% cheaper
Prompts over 272,000 tokens Cached input $0.02 $0.01 GLM 4.7 FlashX is 50% cheaper
Prompts over 272,000 tokens Output $0.75 $0.40 GLM 4.7 FlashX is 47% cheaper

Blended: GPT-6 Luna $0.20, GLM 4.7 FlashX $0.1525 per 1M โ€” one rate at 3:1 input to output, for comparing two models at a glance.

GPT-6 Luna priced by OpenAI, GLM 4.7 FlashX by Z.ai. Standard tier, pay-as-you-go.

Other pricing tiers

Only GPT-6 Luna lists Batch, Flex and Priority, at up to 50% off its own standard rate; GLM 4.7 FlashX offers none of them.

Every tier either model sells is here, so this is the whole price sheet. A saving is measured against that model's own standard rate โ€” how we read tiers.

Tier GPT-6 Luna GLM 4.7 FlashX Cheaper
Batch only on GPT-6 Luna $0.05 in / $0.25 out 50% off Standard Not offered —
Flex only on GPT-6 Luna $0.05 in / $0.25 out 50% off Standard Not offered —
Priority only on GPT-6 Luna $0.20 in / $1.00 out Not offered —

Benchmarks

GPT-6 Luna leads on all 5 figures, furthest ahead on long-context reasoning (AA-LCR), by 41.6 points.

Expert-exam performance

Humanity's Last Exam โ€” Expert-level questions across many academic domains

GPT-6 Luna 38.5
GLM 4.7 FlashX 7.6

GPT-6 Luna ahead by 30.9

Long-context reasoning

AA-LCR โ€” Long-context reasoning across large inputs

GPT-6 Luna 83.3
GLM 4.7 FlashX 41.7

GPT-6 Luna ahead by 41.6

Physics research reasoning

CritPt โ€” Unpublished physics research reasoning problems

GPT-6 Luna 19.4
GLM 4.7 FlashX 0.3

GPT-6 Luna ahead by 19.1

Factual accuracy

Omniscience Accuracy โ€” Breadth of factual knowledge across domains

GPT-6 Luna 43.8
GLM 4.7 FlashX 16.2

GPT-6 Luna ahead by 27.6

Answer reliability

Omniscience Non-Hallucination โ€” How reliably the model avoids fabricated answers

GPT-6 Luna 23.3
GLM 4.7 FlashX 6.1

GPT-6 Luna ahead by 17.2

Each figure is that model's best published run, with the effort level named beside it โ€” where these scores come from.

Specifications

GPT-6 Luna holds 847,248 more input tokens in one request. GPT-6 Luna also takes file and image. Only GPT-6 Luna lists web search, code execution and computer use; only GLM 4.7 FlashX lists open weights.

Specification GPT-6 Luna GLM 4.7 FlashX
Context window 1,050,000 tokens 202,752 tokens
Max output 128,000 tokens 128,000 tokens
Takes in Text, image, file Text
Puts out Text Text
Knowledge cutoff 18 May 2026 โ€”
Released 22 Sep 2026 19 Jan 2026
Status Active Active
Sold by OpenAI Z.ai
Tool use Yes Yes
Structured outputs Yes Yes
Web search Yes Not listed
Prompt caching Yes Yes
Code execution Yes Not listed
Computer use Yes Not listed
Open weights Not listed Yes
Reasoning Yes Yes

FAQs

Is GPT-6 Luna cheaper than GLM 4.7 FlashX?
No โ€” the other way round. A 1,000-token prompt with a 500-token reply costs $0.0004 on GPT-6 Luna and $0.0003 on GLM 4.7 FlashX, and the same model is cheaper on every workload on this page.
How much do GPT-6 Luna and GLM 4.7 FlashX cost per 1M tokens?
GPT-6 Luna costs $0.10 for input and $0.50 for output. GLM 4.7 FlashX costs $0.07 and $0.40. Standard pay-as-you-go rates, as of 22 Sep 2026.
Which is better, GPT-6 Luna or GLM 4.7 FlashX?
On published benchmarks, GPT-6 Luna. It leads on 5 of the 5 figures both models report, including expert-exam performance (Humanity's Last Exam), where it scores 38.5 against 7.6.
Does GPT-6 Luna or GLM 4.7 FlashX offer batch pricing?
GPT-6 Luna only. Its batch input costs $0.05, 50% off its standard rate, and GLM 4.7 FlashX lists no batch tier at all.
Does GPT-6 Luna or GLM 4.7 FlashX have a bigger context window?
GPT-6 Luna, at 1,050,000 tokens against 202,752. That is 847,248 more input tokens in a single request.

Explore these models

GPT-6 Luna

OpenAI ยท 1,050,000-token context

Pricing · Calculator · Specifications

GLM 4.7 FlashX

Z.ai ยท 202,752-token context

Pricing · Calculator · Specifications

Other comparisons