GPT-5.4 Nano vs GLM 4.6

GPT-5.4 Nano is the cheaper of the two, though not evenly: its input costs 67% less than GLM 4.6's, $0.20 per 1M against $0.60, while its output is only 43% cheaper, $1.25 against $2.20. Caching widens the gap: over a thousand agent loops, where most of each prompt comes from cache, that is $37.50 against $120.80.

GPT-5.4 Nano also scores higher on coding ability (Coding Index), 56.1 (xhigh) against 45.8 (Reasoning), so it is cheaper and better at once โ€” about as easy as this choice gets.

Standard pay-as-you-go prices, as of 30 Sep 2025. Last updated 30 Sep 2026.

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

๐Ÿ”’ We respect your privacy. Unsubscribe anytime.

Cheaper overall

GPT-5.4 Nano

54% lower blended rate

Higher benchmark scores

GPT-5.4 Nano

leads on 9 of 11 figures

Bigger context window

GPT-5.4 Nano

400,000 vs 204,800 tokens

Cheaper cached input

GPT-5.4 Nano

$0.02 vs $0.11, 82% less

Where each one wins

GPT-5.4 Nano

  • Cheaper input tokens $0.20 vs $0.60, 67% less
  • Cheaper output tokens $1.25 vs $2.20, 43% less
  • Cheaper cached input $0.02 vs $0.11, 82% less
  • Cheaper on all four workloads
  • Larger context window 400,000 tokens
  • Longer maximum output 128,000 tokens
  • Higher coding ability (Coding Index) 56.1 vs 45.8
  • Higher graduate-level science (GPQA Diamond) 81.7 vs 78
  • Higher expert-exam performance (Humanity's Last Exam) 28.3 vs 14.5
  • Ahead on 6 more benchmark figures
  • Offers Batch and Flex pricing not listed for GLM 4.6

GLM 4.6

  • Higher support-agent performance (ฯ„ยฒ-Bench Telecom) 76.9 vs 76
  • Higher factual accuracy (Omniscience Accuracy) 26.9 vs 25.7

What four workloads cost

One run, and the same run a thousand times.

GPT-5.4 Nano is cheaper on all four workloads, by much the same margin each time โ€” about 2.6 times, from $0.0008 against $0.0017 on a chat turn to $0.0375 against $0.1208 on a long document.

Workload GPT-5.4 Nano GLM 4.6 ร—1,000 Difference
Chat turn 1,000 in / 500 out $0.0008 $0.0017 $0.8250 / $1.70 GPT-5.4 Nano is 51% cheaper
RAG answer 10,000 in / 800 out $0.0030 $0.0078 $3.00 / $7.76 GPT-5.4 Nano is 61% cheaper
Agent loop 32,000 in / 700 out ร— 20 calls $0.0375 $0.1208 $37.50 / $120.80 GPT-5.4 Nano is 69% cheaper
Long document 400,000 in / 3,000 out GLM 4.6's context window holds 204,800 tokens, so a 400,000-token prompt does not fit.

Price your own token counts for these two โ†’

Price per 1M tokens

Above 272,000 input tokens the rates change, and the new rate applies to the whole request, not just the tokens past that point. It does not change which of the two is cheaper.

Context band Price GPT-5.4 Nano GLM 4.6 Difference
Prompts up to 272,000 tokens Input $0.20 $0.60 GPT-5.4 Nano is 67% cheaper
Prompts up to 272,000 tokens Cached input $0.02 $0.11 GPT-5.4 Nano is 82% cheaper
Prompts up to 272,000 tokens Output $1.25 $2.20 GPT-5.4 Nano is 43% cheaper
Prompts over 272,000 tokens Input $0.20 $0.60 GPT-5.4 Nano is 67% cheaper
Prompts over 272,000 tokens Cached input $0.02 $0.11 GPT-5.4 Nano is 82% cheaper
Prompts over 272,000 tokens Output $1.25 $2.20 GPT-5.4 Nano is 43% cheaper

Blended: GPT-5.4 Nano $0.4625, GLM 4.6 $1.00 per 1M โ€” one rate at 3:1 input to output, for comparing two models at a glance.

GPT-5.4 Nano priced by OpenAI, GLM 4.6 by Z.ai. Standard tier, pay-as-you-go.

Other pricing tiers

Only GPT-5.4 Nano lists Batch and Flex, at up to 50% off its own standard rate; GLM 4.6 offers none of them.

Every tier either model sells is here, so this is the whole price sheet. A saving is measured against that model's own standard rate โ€” how we read tiers.

Tier GPT-5.4 Nano GLM 4.6 Cheaper
Batch only on GPT-5.4 Nano $0.10 in / $0.625 out 50% off Standard Not offered —
Flex only on GPT-5.4 Nano $0.10 in / $0.625 out 50% off Standard Not offered —

Benchmarks

They share 11 figures: GPT-5.4 Nano leads on 9, GLM 4.6 on 2. The widest gap is 32.5 points, on instruction following (IFBench).

Coding ability

Coding Index โ€” Coding ability across software-engineering evals

GPT-5.4 Nano 56.1
GLM 4.6 45.8

GPT-5.4 Nano ahead by 10.3

Graduate-level science

GPQA Diamond โ€” Graduate-level questions in biology, chemistry & physics

GPT-5.4 Nano 81.7
GLM 4.6 78

GPT-5.4 Nano ahead by 3.7

Expert-exam performance

Humanity's Last Exam โ€” Expert-level questions across many academic domains

GPT-5.4 Nano 28.3
GLM 4.6 14.5

GPT-5.4 Nano ahead by 13.8

Instruction following

IFBench โ€” Precise following of detailed instructions

GPT-5.4 Nano 75.9
GLM 4.6 43.4

GPT-5.4 Nano ahead by 32.5

Support-agent performance

ฯ„ยฒ-Bench Telecom โ€” Tool-using agent tasks in a telecom support setting

GPT-5.4 Nano 76
GLM 4.6 76.9

GLM 4.6 ahead by 0.9

Long-context reasoning

AA-LCR โ€” Long-context reasoning across large inputs

GPT-5.4 Nano 76.7
GLM 4.6 54

GPT-5.4 Nano ahead by 22.7

Command-line work

Terminal-Bench Hard โ€” Complex command-line and terminal workflows

GPT-5.4 Nano 42.4
GLM 4.6 28.8

GPT-5.4 Nano ahead by 13.6

Physics research reasoning

CritPt โ€” Unpublished physics research reasoning problems

GPT-5.4 Nano 9.3
GLM 4.6 1.1

GPT-5.4 Nano ahead by 8.2

Real-world task performance

GDPval โ€” Economically valuable, real-world knowledge work

GPT-5.4 Nano 21.8
GLM 4.6 12.3

GPT-5.4 Nano ahead by 9.5

Factual accuracy

Omniscience Accuracy โ€” Breadth of factual knowledge across domains

GPT-5.4 Nano 25.7
GLM 4.6 26.9

GLM 4.6 ahead by 1.2

Answer reliability

Omniscience Non-Hallucination โ€” How reliably the model avoids fabricated answers

GPT-5.4 Nano 48.9
GLM 4.6 32.4

GPT-5.4 Nano ahead by 16.5

Each figure is that model's best published run, with the effort level named beside it โ€” where these scores come from.

Specifications

GPT-5.4 Nano holds 195,200 more input tokens in one request. GPT-5.4 Nano also takes file and image.

Specification GPT-5.4 Nano GLM 4.6
Context window 400,000 tokens 204,800 tokens
Max output 128,000 tokens 16,384 tokens
Takes in Text, image, file Text
Puts out Text Text
Knowledge cutoff 31 Aug 2025 31 Mar 2025
Released 17 Mar 2026 30 Sep 2025
Status Active Active
Sold by OpenAI, Perplexity Z.ai

FAQs

Is GPT-5.4 Nano cheaper than GLM 4.6?
Yes. A 1,000-token prompt with a 500-token reply costs $0.0008 on GPT-5.4 Nano and $0.0017 on GLM 4.6, and the same model is cheaper on every workload on this page.
How much do GPT-5.4 Nano and GLM 4.6 cost per 1M tokens?
GPT-5.4 Nano costs $0.20 for input and $1.25 for output. GLM 4.6 costs $0.60 and $2.20. Standard pay-as-you-go rates, as of 30 Sep 2025.
Which is better, GPT-5.4 Nano or GLM 4.6?
On published benchmarks, GPT-5.4 Nano. It leads on 9 of the 11 figures both models report, including coding ability (Coding Index), where it scores 56.1 against 45.8.
Does GPT-5.4 Nano or GLM 4.6 offer batch pricing?
GPT-5.4 Nano only. Its batch input costs $0.10, 50% off its standard rate, and GLM 4.6 lists no batch tier at all.
Does GPT-5.4 Nano or GLM 4.6 have a bigger context window?
GPT-5.4 Nano, at 400,000 tokens against 204,800. That is 195,200 more input tokens in a single request.

Explore these models

GPT-5.4 Nano

OpenAI ยท 400,000-token context

Pricing · Calculator · Specifications

GLM 4.6

Z.ai ยท 204,800-token context

Pricing · Calculator · Specifications

Other comparisons