Gemini 3.8 Flash Pricing - Cost Calculator

Gemini 3.8 Flash is Google’s most intelligent Flash model for long-horizon coding, autonomous agents, complex reasoning, and demanding enterprise workflows.

Gemini 3.8 Flash, released 2 September 2026, is a multimodal model from Google. It pairs a 1,048,576-token context window with up to 65,536 tokens of output, accepting text, audio, image, video and file and returning text. Step-by-step reasoning is supported. It costs $0.75 per million input tokens through Perplexity.

Last updated Sep 17, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

πŸ”’ We respect your privacy. Unsubscribe anytime.

Model Details

Released
Sep 2, 2026
Context Length
1,048,576
Max Output
65,536
Modalities
Text Audio Image Video File Text
Capabilities
Tool use Structured outputs Prompt caching Code execution Computer use

Where is Gemini 3.8 Flash a perfect fit?

Gemini 3.8 Flash is Google’s most capable Flash model, combining advanced reasoning, coding, multimodal understanding, iterative tool use, and autonomous agent capabilities with Flash-level speed and cost efficiency.
- Long-horizon software engineering
- Autonomous coding agents
- Multi-file codebase refactoring
- Complex multi-step reasoning
- Enterprise automation
- Financial and legal analysis
- Agentic workflows with iterative tool calls
- Long-context document and data analysis
- Browser/computer-use agents
- High-volume production workloads
Gemini 3.8 Flash is specifically optimized for long-running coding and autonomous agent workflows rather than merely maximizing low-latency throughput. Google reports substantial gains over 3.7 Flash in software engineering, agentic tasks, and specialized professional reasoning.

Quick Model Estimate

(USD 0.7500 per 1M tokens)
(USD 3.7500 per 1M tokens)

Your Gemini 3.8 Flash Cost Estimate

πŸ’° Total Cost

β€”

for 1000 input + 1000 output tokens

πŸ“₯ Input (1000 Γ— $0.750000) β€”
πŸ“€ Output (1000 Γ— $3.750000) β€”

Cost Breakdown

πŸ“₯ Input πŸ“€ Output

Prices updated daily from official provider data.

Pricing

Provider ↕
Modality ↕
Service Tier ↕
Input Price
(per 1M tokens)
↕
Output Price
(per 1M tokens)
↕
Cached Input
(per 1M tokens)
↕
Context Size ↕
View
Perplexity LogoPerplexityTextStandard$0.7500$3.7500$0.07501,048,576 tokens→

FAQs about Gemini 3.8 Flash

How much does Gemini 3.8 Flash cost per 1M tokens?

Gemini 3.8 Flash costs $0.75 per million input tokens, and $3.75 per million output tokens.

What does a typical workload cost with Gemini 3.8 Flash?

1,000 requests of 2,000 input and 500 output tokens each β€” 2,000,000 input and 500,000 output tokens in total β€” costs $3.375 with Gemini 3.8 Flash at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use Gemini 3.8 Flash?

Gemini 3.8 Flash is available through Perplexity, from $0.75 per million input tokens.

What is Gemini 3.8 Flash's context window?

Gemini 3.8 Flash has a context window of 1,048,576 tokens, and returns up to 65,536 tokens in a single response.

What input types does Gemini 3.8 Flash support?

Gemini 3.8 Flash accepts text, audio, image, video and file input, and returns text.