Model Details
Where is gemini-3.5-flash-lite a perfect fit?
- High-volume document parsing, extraction, and structured JSON generation
- Autonomous subagents handling inexpensive portions of larger agentic workflows
- Classification, routing, tagging, summarization, and data transformation at scale
- Multimodal processing across text, images, video, audio, and PDFs
- Computer-use agents performing UI automation
- Function calling, code execution, search, and tool-driven workflows
- High-throughput applications where latency and cost matter more than maximum reasoning depth
- Migration from Gemini 3.1 Flash-Lite or Gemini 2.5 Flash when stronger reasoning and multimodal performance are useful
Quick Model Estimate
Your GPT-5 Cost Estimate
💰 Total Cost
USD 3.00
for 1000 input + 1000 output tokens
Cost Breakdown
Prices updated daily from official provider data.
Pricing
|
Provider
↕
|
Modality
↕
|
Service Tier
↕
|
Input Price
(per 1M tokens) ↕ |
Output Price
(per 1M tokens) ↕ |
Cached Input
(per 1M tokens) ↕ |
Context Size
↕
|
View
|
|---|---|---|---|---|---|---|---|
Perplexity | Text | Standard | $0.3000 | $2.5000 | $0.0300 | 1,048,576 tokens | → |
FAQs about gemini-3.5-flash-lite
How much does gemini-3.5-flash-lite cost per 1M tokens?
gemini-3.5-flash-lite costs $0.30 per million input tokens, and $2.50 per million output tokens.
What does a typical workload cost with gemini-3.5-flash-lite?
1,000 requests of 2,000 input and 500 output tokens each — 2,000,000 input and 500,000 output tokens in total — costs $1.85 with gemini-3.5-flash-lite at its lowest rates. Use the calculator on this page for your own volumes.
Where can I use gemini-3.5-flash-lite?
gemini-3.5-flash-lite is available through Perplexity, from $0.30 per million input tokens.
What is gemini-3.5-flash-lite's context window?
gemini-3.5-flash-lite has a context window of 1,048,576 tokens, and returns up to 65,536 tokens in a single response.
