Model Details
Where is Gemini 3.7 Flash a perfect fit?
- Advanced software engineering and code generation
- Autonomous coding agents
- Multi-step agentic workflows and tool calling
- Web and application development
- Codebase research, debugging, and verification
- Multimodal reasoning over images, audio, and video
- Long-context document and knowledge-work applications
- Production workloads requiring a strong balance of intelligence, speed, and cost
Quick Model Estimate
Your Gemini 3.7 Flash Cost Estimate
💰 Total Cost
—
for 1000 input + 1000 output tokens
Cost Breakdown
Prices are watched for changes and checked against the provider's own page.
Pricing
|
Provider
↕
|
Modality
↕
|
Service Tier
↕
|
Input Price
(per 1M tokens) ↕ |
Output Price
(per 1M tokens) ↕ |
Cached Input
(per 1M tokens) ↕ |
Context Size
↕
|
View
|
|---|---|---|---|---|---|---|---|
Perplexity | Text | Standard | $0.7500 | $3.7500 | $0.0750 | 1,048,576 tokens | → |
Price changes
Every recorded change to Gemini 3.7 Flash’s price, most recent first.
-
Perplexity raised the price of Gemini 3.7 Flash, effective 4 September 2026.
Input $0.375 → $0.75 per 1M tokens (+100%), and 2 other price fields.
Per-configuration detail on Perplexity →
FAQs about Gemini 3.7 Flash
How much does Gemini 3.7 Flash cost per 1M tokens?
Gemini 3.7 Flash costs $0.75 per million input tokens, and $3.75 per million output tokens.
What does a typical workload cost with Gemini 3.7 Flash?
1,000 requests of 2,000 input and 500 output tokens each — 2,000,000 input and 500,000 output tokens in total — costs $3.375 with Gemini 3.7 Flash at its lowest rates. Use the calculator on this page for your own volumes.
Has Gemini 3.7 Flash's price changed?
Yes, once: Perplexity raised it on 4 September 2026, from $0.375 to $0.75 per million input tokens.
Where can I use Gemini 3.7 Flash?
Gemini 3.7 Flash is available through Perplexity, from $0.75 per million input tokens.
What is Gemini 3.7 Flash's context window?
Gemini 3.7 Flash has a context window of 1,048,576 tokens, and returns up to 65,536 tokens in a single response.
