gemini-3.5-flash Pricing - Cost Calculator

Track gemini-3.5-flash pricing with real-time alerts, compare providers, and monitor cost changes across advanced agentic, coding, and multimodal AI workloads.

Released 19 May 2026, gemini-3.5-flash is a multimodal model from Google. It works across a 1,048,576-token context window and returns up to 65,536 tokens. Input is priced at $1.50 per million tokens through Perplexity.

Last updated Sep 2, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

🔒 We respect your privacy. Unsubscribe anytime.

Model Details

Released
May 19, 2026
Context Length
1,048,576
Max Output
65,536

Where is gemini-3.5-flash a perfect fit?

Gemini 3.5 Flash is Google's production Flash model designed for sustained frontier-level performance while retaining the speed and economics needed for scaled deployment. Google particularly emphasizes agentic execution, coding, multi-step workflows, and long-horizon tasks.
- Complex coding, debugging, and iterative software-development workflows
- Multi-step autonomous agents and sub-agent orchestration
- Long-horizon tasks involving repeated reasoning and tool calls
- Computer-use agents interacting with browser, mobile, and desktop interfaces
- Multimodal analysis across text, images, video, audio, and PDFs
- Search-grounded and Maps-grounded applications
- Function calling, structured data extraction, and code execution
- Production applications needing stronger intelligence while retaining Flash-class speed

Quick Model Estimate

(USD 1.5000 per 1M tokens)
(USD 9.0000 per 1M tokens)

Your GPT-5 Cost Estimate

💰 Total Cost

USD 3.00

for 1000 input + 1000 output tokens

📥 Input (1000 × $1.500000) USD 1.5000
📤 Output (1000 × $9.000000) USD 1.5000

Cost Breakdown

📥 Input 50% 📤 Output 50%

Prices updated daily from official provider data.

Pricing

Provider
Modality
Service Tier
Input Price
(per 1M tokens)
Output Price
(per 1M tokens)
Cached Input
(per 1M tokens)
Context Size
View
Perplexity LogoPerplexityTextStandard$1.5000$9.0000$0.15001,048,576 tokens

FAQs about gemini-3.5-flash

How much does gemini-3.5-flash cost per 1M tokens?

gemini-3.5-flash costs $1.50 per million input tokens, and $9.00 per million output tokens.

What does a typical workload cost with gemini-3.5-flash?

1,000 requests of 2,000 input and 500 output tokens each — 2,000,000 input and 500,000 output tokens in total — costs $7.50 with gemini-3.5-flash at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use gemini-3.5-flash?

gemini-3.5-flash is available through Perplexity, from $1.50 per million input tokens.

What is gemini-3.5-flash's context window?

gemini-3.5-flash has a context window of 1,048,576 tokens, and returns up to 65,536 tokens in a single response.