gemini-3.6-flash Pricing - Cost Calculator

Track gemini-3.6-flash pricing with real-time alerts, compare providers, and monitor cost changes across advanced agentic, coding, and multimodal AI workloads.

Released 21 July 2026, gemini-3.6-flash is a multimodal model from Google. It pairs a 1,048,576-token context window with up to 65,536 tokens of output. Perplexity charges $1.50 per million input tokens.

Last updated Sep 2, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

🔒 We respect your privacy. Unsubscribe anytime.

Model Details

Released
Jul 21, 2026
Context Length
1,048,576
Max Output
65,536

Where is gemini-3.6-flash a perfect fit?

Gemini 3.6 Flash is Google's production-ready Flash model designed for frontier-level intelligence at higher speed and lower cost. It particularly targets rapid agentic loops, coding, multimodal understanding, and spatial reasoning.
- Complex coding, debugging, and iterative software-development workflows
- Autonomous agents requiring repeated reasoning and tool calls
- Rapid agentic loops involving complex coding cycles
- Computer-use agents performing browser and UI automation
- Multimodal analysis across text, images, video, audio, and PDFs
- Spatial reasoning, chart interpretation, blueprints, and visual layouts
- Search-grounded applications requiring current information
- Production workloads where reducing reasoning tokens, tool calls, and overall inference cost matters

Quick Model Estimate

(USD 1.5000 per 1M tokens)
(USD 7.5000 per 1M tokens)

Your GPT-5 Cost Estimate

💰 Total Cost

USD 3.00

for 1000 input + 1000 output tokens

📥 Input (1000 × $1.500000) USD 1.5000
📤 Output (1000 × $7.500000) USD 1.5000

Cost Breakdown

📥 Input 50% 📤 Output 50%

Prices updated daily from official provider data.

Pricing

Provider
Modality
Service Tier
Input Price
(per 1M tokens)
Output Price
(per 1M tokens)
Cached Input
(per 1M tokens)
Context Size
View
Perplexity LogoPerplexityTextStandard$1.5000$7.5000$0.15001,048,576 tokens

FAQs about gemini-3.6-flash

How much does gemini-3.6-flash cost per 1M tokens?

gemini-3.6-flash costs $1.50 per million input tokens, and $7.50 per million output tokens.

What does a typical workload cost with gemini-3.6-flash?

1,000 requests of 2,000 input and 500 output tokens each — 2,000,000 input and 500,000 output tokens in total — costs $6.75 with gemini-3.6-flash at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use gemini-3.6-flash?

gemini-3.6-flash is available through Perplexity, from $1.50 per million input tokens.

What is gemini-3.6-flash's context window?

gemini-3.6-flash has a context window of 1,048,576 tokens, and returns up to 65,536 tokens in a single response.