Model Details
Where is Gemini 3.8 Flash a perfect fit?
- Long-horizon software engineering
- Autonomous coding agents
- Multi-file codebase refactoring
- Complex multi-step reasoning
- Enterprise automation
- Financial and legal analysis
- Agentic workflows with iterative tool calls
- Long-context document and data analysis
- Browser/computer-use agents
- High-volume production workloads
Gemini 3.8 Flash is specifically optimized for long-running coding and autonomous agent workflows rather than merely maximizing low-latency throughput. Google reports substantial gains over 3.7 Flash in software engineering, agentic tasks, and specialized professional reasoning.
Quick Model Estimate
Your Gemini 3.8 Flash Cost Estimate
π° Total Cost
β
for 1000 input + 1000 output tokens
Cost Breakdown
Prices updated daily from official provider data.
Pricing
|
Provider
β
|
Modality
β
|
Service Tier
β
|
Input Price
(per 1M tokens) β |
Output Price
(per 1M tokens) β |
Cached Input
(per 1M tokens) β |
Context Size
β
|
View
|
|---|---|---|---|---|---|---|---|
Perplexity | Text | Standard | $0.7500 | $3.7500 | $0.0750 | 1,048,576 tokens | β |
FAQs about Gemini 3.8 Flash
How much does Gemini 3.8 Flash cost per 1M tokens?
Gemini 3.8 Flash costs $0.75 per million input tokens, and $3.75 per million output tokens.
What does a typical workload cost with Gemini 3.8 Flash?
1,000 requests of 2,000 input and 500 output tokens each β 2,000,000 input and 500,000 output tokens in total β costs $3.375 with Gemini 3.8 Flash at its lowest rates. Use the calculator on this page for your own volumes.
Where can I use Gemini 3.8 Flash?
Gemini 3.8 Flash is available through Perplexity, from $0.75 per million input tokens.
What is Gemini 3.8 Flash's context window?
Gemini 3.8 Flash has a context window of 1,048,576 tokens, and returns up to 65,536 tokens in a single response.
What input types does Gemini 3.8 Flash support?
Gemini 3.8 Flash accepts text, audio, image, video and file input, and returns text.
