gemini-3.5-flash-lite Pricing - Cost Calculator

Track gemini-3.5-flash-lite pricing with real-time alerts, compare providers, and monitor cost changes across high-volume, cost-efficient multimodal AI workloads at scale.

Released 21 July 2026, gemini-3.5-flash-lite is a multimodal model from Google. It pairs a 1,048,576-token context window with up to 65,536 tokens of output. Input is priced at $0.30 per million tokens through Perplexity.

Last updated Sep 2, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

🔒 We respect your privacy. Unsubscribe anytime.

Model Details

Released
Jul 21, 2026
Context Length
1,048,576
Max Output
65,536

Where is gemini-3.5-flash-lite a perfect fit?

Gemini 3.5 Flash-Lite is Google's fastest and lowest-cost model in the Gemini 3.5 family, optimized for high-throughput execution, subagent tasks, document parsing, and workloads where latency and API cost are major constraints.
- High-volume document parsing, extraction, and structured JSON generation
- Autonomous subagents handling inexpensive portions of larger agentic workflows
- Classification, routing, tagging, summarization, and data transformation at scale
- Multimodal processing across text, images, video, audio, and PDFs
- Computer-use agents performing UI automation
- Function calling, code execution, search, and tool-driven workflows
- High-throughput applications where latency and cost matter more than maximum reasoning depth
- Migration from Gemini 3.1 Flash-Lite or Gemini 2.5 Flash when stronger reasoning and multimodal performance are useful

Quick Model Estimate

(USD 0.3000 per 1M tokens)
(USD 2.5000 per 1M tokens)

Your GPT-5 Cost Estimate

💰 Total Cost

USD 3.00

for 1000 input + 1000 output tokens

📥 Input (1000 × $0.300000) USD 1.5000
📤 Output (1000 × $2.500000) USD 1.5000

Cost Breakdown

📥 Input 50% 📤 Output 50%

Prices updated daily from official provider data.

Pricing

Provider
Modality
Service Tier
Input Price
(per 1M tokens)
Output Price
(per 1M tokens)
Cached Input
(per 1M tokens)
Context Size
View
Perplexity LogoPerplexityTextStandard$0.3000$2.5000$0.03001,048,576 tokens

FAQs about gemini-3.5-flash-lite

How much does gemini-3.5-flash-lite cost per 1M tokens?

gemini-3.5-flash-lite costs $0.30 per million input tokens, and $2.50 per million output tokens.

What does a typical workload cost with gemini-3.5-flash-lite?

1,000 requests of 2,000 input and 500 output tokens each — 2,000,000 input and 500,000 output tokens in total — costs $1.85 with gemini-3.5-flash-lite at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use gemini-3.5-flash-lite?

gemini-3.5-flash-lite is available through Perplexity, from $0.30 per million input tokens.

What is gemini-3.5-flash-lite's context window?

gemini-3.5-flash-lite has a context window of 1,048,576 tokens, and returns up to 65,536 tokens in a single response.