GPT-6 Luna vs GLM 5.3 Flash

GPT-6 Luna and GLM 5.3 Flash charge the same $0.50 per 1M for output. The whole difference is on the input side, where GPT-6 Luna asks $0.10 against $0.15, 33% less. They stay level up to 272,000 input tokens; past that, GLM 5.3 Flash is the cheaper of the two.

On overall intelligence (Intelligence Index), GLM 5.3 Flash is ahead, 41.9 against 37.3 (max). With the costs split between them, that is the tie-breaker.

Standard pay-as-you-go prices, as of 22 Sep 2026. Last updated 27 Sep 2026.

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

๐Ÿ”’ We respect your privacy. Unsubscribe anytime.

Cheaper overall

GPT-6 Luna

16% lower blended rate

Cheaper on long prompts

GLM 5.3 Flash

above 272,000 input tokens

Bigger context window

GPT-6 Luna

1,050,000 vs 1,000,000 tokens

Cheaper cached input

GPT-6 Luna

$0.01 vs $0.03, 67% less

Where each one wins

GPT-6 Luna

  • Cheaper input tokens $0.10 vs $0.15, 33% less
  • Cheaper cached input $0.01 vs $0.03, 67% less
  • Larger context window 1,050,000 tokens
  • Higher long-context reasoning (AA-LCR) 83.3 vs 80
  • Higher scientific coding (SciCode) 54.6 vs 51.6
  • Higher physics research reasoning (CritPt) 19.4 vs 15.4
  • Ahead on 1 more benchmark figure
  • Offers Batch, Flex and Priority pricing not listed for GLM 5.3 Flash

GLM 5.3 Flash

  • Cheaper on long prompts above 272,000 input tokens
  • Higher overall intelligence (Intelligence Index) 41.9 vs 37.3
  • Higher expert-exam performance (Humanity's Last Exam) 39.9 vs 38.5
  • Higher real-world task performance (GDPval) 57.7 vs 43.4
  • Ahead on 1 more benchmark figure

What four workloads cost

One run, and the same run a thousand times.

The gap opens on long prompts: GLM 5.3 Flash costs $0.0615 against $0.0823, a difference of $20.75 over a thousand runs.

Workload GPT-6 Luna GLM 5.3 Flash ร—1,000 Difference
Chat turn 1,000 in / 500 out $0.0004 $0.0004 $0.3500 / $0.4000 GPT-6 Luna is 13% cheaper
RAG answer 10,000 in / 800 out $0.0014 $0.0019 $1.40 / $1.90 GPT-6 Luna is 26% cheaper
Agent loop 32,000 in / 700 out ร— 20 calls $0.0170 $0.0310 $17.00 / $31.00 GPT-6 Luna is 45% cheaper
Long document 400,000 in / 3,000 out $0.0823 $0.0615 $82.25 / $61.50 GLM 5.3 Flash is 25% cheaper

Price your own token counts for these two โ†’

Price per 1M tokens

Above 272,000 input tokens the rates change, and the new rate applies to the whole request, not just the tokens past that point: GPT-6 Luna's input goes from $0.1000 to $0.2000. That is where the cheaper model changes.

Context band Price GPT-6 Luna GLM 5.3 Flash Difference
Prompts up to 272,000 tokens Input $0.10 $0.15 GPT-6 Luna is 33% cheaper
Prompts up to 272,000 tokens Cached input $0.01 $0.03 GPT-6 Luna is 67% cheaper
Prompts up to 272,000 tokens Output $0.50 $0.50 Same
Prompts over 272,000 tokens Input $0.20 $0.15 GLM 5.3 Flash is 25% cheaper
Prompts over 272,000 tokens Cached input $0.02 $0.03 GPT-6 Luna is 33% cheaper
Prompts over 272,000 tokens Output $0.75 $0.50 GLM 5.3 Flash is 33% cheaper

Blended: GPT-6 Luna $0.20, GLM 5.3 Flash $0.2375 per 1M โ€” one rate at 3:1 input to output, for comparing two models at a glance.

GPT-6 Luna priced by OpenAI, GLM 5.3 Flash by Z.ai. Standard tier, pay-as-you-go.

Other pricing tiers

Only GPT-6 Luna lists Batch, Flex and Priority, at up to 50% off its own standard rate; GLM 5.3 Flash offers none of them.

Every tier either model sells is here, so this is the whole price sheet. A saving is measured against that model's own standard rate โ€” how we read tiers.

Tier GPT-6 Luna GLM 5.3 Flash Cheaper
Batch only on GPT-6 Luna $0.05 in / $0.25 out 50% off Standard Not offered —
Flex only on GPT-6 Luna $0.05 in / $0.25 out 50% off Standard Not offered —
Priority only on GPT-6 Luna $0.20 in / $1.00 out Not offered —

Benchmarks

They share 8 figures: GPT-6 Luna leads on 4, GLM 5.3 Flash on 4. The widest gap is 49.1 points, on answer reliability (Omniscience Non-Hallucination).

Overall intelligence

Intelligence Index โ€” Overall intelligence across reasoning, knowledge & math evals

GPT-6 Luna 37.3
GLM 5.3 Flash 41.9

GLM 5.3 Flash ahead by 4.6

Expert-exam performance

Humanity's Last Exam โ€” Expert-level questions across many academic domains

GPT-6 Luna 38.5
GLM 5.3 Flash 39.9

GLM 5.3 Flash ahead by 1.4

Long-context reasoning

AA-LCR โ€” Long-context reasoning across large inputs

GPT-6 Luna 83.3
GLM 5.3 Flash 80

GPT-6 Luna ahead by 3.3

Scientific coding

SciCode โ€” Research-level scientific coding tasks

GPT-6 Luna 54.6
GLM 5.3 Flash 51.6

GPT-6 Luna ahead by 3

Physics research reasoning

CritPt โ€” Unpublished physics research reasoning problems

GPT-6 Luna 19.4
GLM 5.3 Flash 15.4

GPT-6 Luna ahead by 4

Real-world task performance

GDPval โ€” Economically valuable, real-world knowledge work

GPT-6 Luna 43.4
GLM 5.3 Flash 57.7

GLM 5.3 Flash ahead by 14.3

Factual accuracy

Omniscience Accuracy โ€” Breadth of factual knowledge across domains

GPT-6 Luna 43.8
GLM 5.3 Flash 27.5

GPT-6 Luna ahead by 16.3

Answer reliability

Omniscience Non-Hallucination โ€” How reliably the model avoids fabricated answers

GPT-6 Luna 23.3
GLM 5.3 Flash 72.4

GLM 5.3 Flash ahead by 49.1

Each figure is that model's best published run, with the effort level named beside it โ€” where these scores come from.

Specifications

GPT-6 Luna holds 50,000 more input tokens in one request. GLM 5.3 Flash also takes video. Only GPT-6 Luna lists web search; only GLM 5.3 Flash lists open weights.

Specification GPT-6 Luna GLM 5.3 Flash
Context window 1,050,000 tokens 1,000,000 tokens
Max output 128,000 tokens 128,000 tokens
Takes in Text, image, file Text, image, video, file
Puts out Text Text
Knowledge cutoff 18 May 2026 โ€”
Released 22 Sep 2026 26 Aug 2026
Status Active Active
Sold by OpenAI Perplexity, Z.ai
Tool use Yes Yes
Structured outputs Yes Yes
Web search Yes Not listed
Prompt caching Yes Yes
Code execution Yes Yes
Computer use Yes Yes
Open weights Not listed Yes
Reasoning Yes Yes

FAQs

Is GPT-6 Luna cheaper than GLM 5.3 Flash?
Yes. A 1,000-token prompt with a 500-token reply costs $0.0004 on GPT-6 Luna and $0.0004 on GLM 5.3 Flash.
How much do GPT-6 Luna and GLM 5.3 Flash cost per 1M tokens?
GPT-6 Luna costs $0.10 for input and $0.50 for output. GLM 5.3 Flash costs $0.15 and $0.50. Standard pay-as-you-go rates, as of 22 Sep 2026.
Which is better, GPT-6 Luna or GLM 5.3 Flash?
On published benchmarks, GLM 5.3 Flash. It leads on 4 of the 8 figures both models report, including overall intelligence (Intelligence Index), where it scores 41.9 against 37.3.
Is GPT-6 Luna or GLM 5.3 Flash cheaper for long prompts?
GLM 5.3 Flash, above 272,000 input tokens. A rate that changes at a context length applies to the whole request, not only the tokens past it, so the switch is sharp rather than gradual.
Does GPT-6 Luna or GLM 5.3 Flash offer batch pricing?
GPT-6 Luna only. Its batch input costs $0.05, 50% off its standard rate, and GLM 5.3 Flash lists no batch tier at all.
Does GPT-6 Luna or GLM 5.3 Flash have a bigger context window?
GPT-6 Luna, at 1,050,000 tokens against 1,000,000. That is 50,000 more input tokens in a single request.

Explore these models

GPT-6 Luna

OpenAI ยท 1,050,000-token context

Pricing · Calculator · Specifications

GLM 5.3 Flash

Z.ai ยท 1,000,000-token context

Pricing · Calculator · Specifications

Other comparisons