Cheaper overall
Grok-4.7
47% lower blended rate
Cheaper on long prompts
GPT-5.4
above 199,999 input tokens
Higher benchmark scores
GPT-5.4
leads on 5 of 8 figures
Bigger context window
GPT-5.4
1,050,000 vs 500,000 tokens
Cheaper cached input
GPT-5.4
$0.25 vs $0.50, 50% less
Where each one wins
Grok-4.7
- Cheaper input tokens $2.00 vs $2.50, 20% less
- Cheaper output tokens $6.00 vs $15.00, 60% less
- Higher real-world task performance (GDPval) 59.8 vs 36.6
- Higher answer reliability (Omniscience Non-Hallucination) 70.7 vs 17.4
- Higher Gamedev Elo 1,308 vs 1,253
GPT-5.4
- Cheaper cached input $0.25 vs $0.50, 50% less
- Cheaper on long prompts above 199,999 input tokens
- Larger context window 1,050,000 tokens
- Higher expert-exam performance (Humanity's Last Exam) 43.7 vs 43.1
- Higher long-context reasoning (AA-LCR) 82 vs 76.7
- Higher physics research reasoning (CritPt) 23.4 vs 17.7
- Ahead on 2 more benchmark figures
- Offers Batch, Flex and Priority pricing not listed for Grok-4.7
What four workloads cost
One run, and the same run a thousand times.
The gap opens on long prompts: Grok-4.7 costs $1.64 against $2.07, a difference of $431.50 over a thousand runs.
| Workload | Grok-4.7 | GPT-5.4 | ร1,000 | Difference |
|---|---|---|---|---|
| Chat turn 1,000 in / 500 out | $0.0050 | $0.0100 | $5.00 / $10.00 | Grok-4.7 is 50% cheaper |
| RAG answer 10,000 in / 800 out | $0.0248 | $0.0370 | $24.80 / $37.00 | Grok-4.7 is 33% cheaper |
| Agent loop 32,000 in / 700 out ร 20 calls | $0.4640 | $0.4600 | $464.00 / $460.00 | GPT-5.4 is 0.9% cheaper |
| Long document 400,000 in / 3,000 out | $1.64 | $2.07 | $1,636.00 / $2,067.50 | Grok-4.7 is 21% cheaper |
Price per 1M tokens
Above 199,999 input tokens the rates change, and the new rate applies to the whole request, not just the tokens past that point: Grok-4.7's input goes from $2.00 to $4.00. That is where the cheaper model changes.
| Context band | Price | Grok-4.7 | GPT-5.4 | Difference |
|---|---|---|---|---|
| Prompts up to 199,999 tokens | Input | $2.00 | $2.50 | Grok-4.7 is 20% cheaper |
| Prompts up to 199,999 tokens | Cached input | $0.50 | $0.25 | GPT-5.4 is 50% cheaper |
| Prompts up to 199,999 tokens | Output | $6.00 | $15.00 | Grok-4.7 is 60% cheaper |
| Prompts up to 200,000 tokens | Input | $4.00 | $2.50 | GPT-5.4 is 38% cheaper |
| Prompts up to 200,000 tokens | Cached input | $1.00 | $0.25 | GPT-5.4 is 75% cheaper |
| Prompts up to 200,000 tokens | Output | $12.00 | $15.00 | Grok-4.7 is 20% cheaper |
| Prompts up to 272,000 tokens | Input | $4.00 | $2.50 | GPT-5.4 is 38% cheaper |
| Prompts up to 272,000 tokens | Cached input | $1.00 | $0.25 | GPT-5.4 is 75% cheaper |
| Prompts up to 272,000 tokens | Output | $12.00 | $15.00 | Grok-4.7 is 20% cheaper |
| Prompts over 272,000 tokens | Input | $4.00 | $5.00 | Grok-4.7 is 20% cheaper |
| Prompts over 272,000 tokens | Cached input | $1.00 | $0.50 | GPT-5.4 is 50% cheaper |
| Prompts over 272,000 tokens | Output | $12.00 | $22.50 | Grok-4.7 is 47% cheaper |
Blended: Grok-4.7 $3.00, GPT-5.4 $5.625 per 1M โ one rate at 3:1 input to output, for comparing two models at a glance.
Grok-4.7 priced by Perplexity, GPT-5.4 by OpenAI. Standard tier, pay-as-you-go.
Other pricing tiers
Only GPT-5.4 lists Batch, Flex and Priority, at up to 50% off its own standard rate; Grok-4.7 offers none of them.
Every tier either model sells is here, so this is the whole price sheet. A saving is measured against that model's own standard rate โ how we read tiers.
| Tier | Grok-4.7 | GPT-5.4 | Cheaper |
|---|---|---|---|
| Batch only on GPT-5.4 | Not offered | $1.25 in / $7.50 out 50% off Standard | — |
| Flex only on GPT-5.4 | Not offered | $1.25 in / $7.50 out 50% off Standard | — |
| Priority only on GPT-5.4 | Not offered | $5.00 in / $30.00 out | — |
Benchmarks
They share 8 figures: Grok-4.7 leads on 3, GPT-5.4 on 5. The widest gap is 55 points, on Gamedev Elo.
Expert-exam performance
Humanity's Last Exam โ Expert-level questions across many academic domains
GPT-5.4 ahead by 0.6
Long-context reasoning
AA-LCR โ Long-context reasoning across large inputs
GPT-5.4 ahead by 5.3
Physics research reasoning
CritPt โ Unpublished physics research reasoning problems
GPT-5.4 ahead by 5.7
Real-world task performance
GDPval โ Economically valuable, real-world knowledge work
Grok-4.7 ahead by 23.2
Factual accuracy
Omniscience Accuracy โ Breadth of factual knowledge across domains
GPT-5.4 ahead by 3.4
Answer reliability
Omniscience Non-Hallucination โ How reliably the model avoids fabricated answers
Grok-4.7 ahead by 53.3
Gamedev Elo
Design Arena head-to-head rating in the gamedev category
Grok-4.7 ahead by 55
Website Elo
Design Arena head-to-head rating in the website category
GPT-5.4 ahead by 4
Each figure is that model's best published run, with the effort level named beside it โ where these scores come from.
Specifications
GPT-5.4 holds 550,000 more input tokens in one request.
| Specification | Grok-4.7 | GPT-5.4 |
|---|---|---|
| Context window | 500,000 tokens | 1,050,000 tokens |
| Max output | โ | 128,000 tokens |
| Takes in | Text, image, file | Text, image, file |
| Puts out | Text | Text |
| Knowledge cutoff | 31 May 2026 | 31 Aug 2025 |
| Released | 21 Sep 2026 | 5 Mar 2026 |
| Status | Active | Active |
| Sold by | Perplexity | OpenAI, Perplexity |
FAQs
- Is Grok-4.7 cheaper than GPT-5.4?
- Yes. A 1,000-token prompt with a 500-token reply costs $0.0050 on Grok-4.7 and $0.0100 on GPT-5.4.
- How much do Grok-4.7 and GPT-5.4 cost per 1M tokens?
- Grok-4.7 costs $2.00 for input and $6.00 for output. GPT-5.4 costs $2.50 and $15.00. Standard pay-as-you-go rates, as of 23 Sep 2026.
- Which is better, Grok-4.7 or GPT-5.4?
- On published benchmarks, GPT-5.4. It leads on 5 of the 8 figures both models report, including expert-exam performance (Humanity's Last Exam), where it scores 43.7 against 43.1.
- Is Grok-4.7 or GPT-5.4 cheaper for long prompts?
- GPT-5.4, above 199,999 input tokens. A rate that changes at a context length applies to the whole request, not only the tokens past it, so the switch is sharp rather than gradual.
- Does Grok-4.7 or GPT-5.4 offer batch pricing?
- GPT-5.4 only. Its batch input costs $1.25, 50% off its standard rate, and Grok-4.7 lists no batch tier at all.
- Does Grok-4.7 or GPT-5.4 have a bigger context window?
- GPT-5.4, at 1,050,000 tokens against 500,000. That is 550,000 more input tokens in a single request.