gemini-3.1-flash-lite Pricing - Cost Calculator

Track gemini-3.1-flash-lite pricing with real-time alerts, compare providers, and monitor cost changes across high-volume, cost-efficient multimodal AI workloads at scale.

Released 7 May 2026, gemini-3.1-flash-lite is a multimodal model from Google. It pairs a 1,048,576-token context window with up to 65,536 tokens of output. Perplexity charges $0.25 per million input tokens.

Last updated Sep 2, 2026

Never miss an AI pricing or launch update

Get notified about new model launches, pricing changes, provider updates, deprecations, and cost-saving insights.

🔒 We respect your privacy. Unsubscribe anytime.

Model Details

Released
May 7, 2026
Context Length
1,048,576
Max Output
65,536

Where is gemini-3.1-flash-lite a perfect fit?

Gemini 3.1 Flash-Lite is Google's low-latency, cost-efficient Gemini model optimized specifically for high-frequency, lightweight workloads at scale.
- High-volume translation of messages, reviews, support tickets, and other content
- Data extraction, classification, tagging, moderation, and transformation
- High-throughput agentic workflows involving tool calling and orchestration
- Multimodal analysis across text, images, video, audio, and PDFs
- Search-grounded applications requiring current external information
- Cost-sensitive applications processing very large request volumes
- Responsive applications where low latency is more important than maximum frontier intelligence
- Code-execution, file-search, structured-output, and function-calling workflows

Quick Model Estimate

(USD 0.2500 per 1M tokens)
(USD 1.5000 per 1M tokens)

Your GPT-5 Cost Estimate

💰 Total Cost

USD 3.00

for 1000 input + 1000 output tokens

📥 Input (1000 × $0.250000) USD 1.5000
📤 Output (1000 × $1.500000) USD 1.5000

Cost Breakdown

📥 Input 50% 📤 Output 50%

Prices updated daily from official provider data.

Pricing

Provider
Modality
Service Tier
Input Price
(per 1M tokens)
Output Price
(per 1M tokens)
Cached Input
(per 1M tokens)
Context Size
View
Perplexity LogoPerplexityTextStandard$0.2500$1.5000$0.02501,048,576 tokens

FAQs about gemini-3.1-flash-lite

How much does gemini-3.1-flash-lite cost per 1M tokens?

gemini-3.1-flash-lite costs $0.25 per million input tokens, and $1.50 per million output tokens.

What does a typical workload cost with gemini-3.1-flash-lite?

1,000 requests of 2,000 input and 500 output tokens each — 2,000,000 input and 500,000 output tokens in total — costs $1.25 with gemini-3.1-flash-lite at its lowest rates. Use the calculator on this page for your own volumes.

Where can I use gemini-3.1-flash-lite?

gemini-3.1-flash-lite is available through Perplexity, from $0.25 per million input tokens.

What is gemini-3.1-flash-lite's context window?

gemini-3.1-flash-lite has a context window of 1,048,576 tokens, and returns up to 65,536 tokens in a single response.