Gemini 3.5 Flash-Lite Pricing & Token Costs (2026)
Per 1M tokens: input $0.30 · output $2.50 · cached $0.030. Context window 1,048,576 tokens with source and verification details.
Reference rates: Developer API · paid tier
USD per million text tokens. Verified 2026-10-07 · Official source
| Mode / prompt size | Input | Cache read | Cache write | Output |
|---|---|---|---|---|
| Standard | $0.3 | $0.03 | Not estimated | $2.5 |
| Batch | $0.15 | $0.02 | Not estimated | $1.25 |
Cache storage: $1 per million tokens per hour, charged separately and excluded from the per-call estimate.
1,048,576 context tokens · max output 65,536. Local token counts are estimates; use the provider’s usage counts for an invoice estimate.
Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.
Gemini Developer API paid-tier text rates (not Vertex AI). Output includes thinking tokens. Cache storage is billed separately by token-hours; grounding and tools cost extra.
TL;DR — Pricing Quick Summary
- ✓Input pricing: $0.30 per 1M tokens ($0.0003 per 1K)
- ✓Output pricing: $2.50 per 1M tokens ($0.0025 per 1K)
- ✓Prompt caching: $0.030 per 1M tokens — save 90% on repeated context
- ✓Context window: 1,048,576 tokens
- ✓Typical monthly cost: $8.00 for 10M input + 2M output across short requests; Developer API · paid tier
- ✓Daily cost example: $0.1400 for 100K tokens (50K in, 50K out)
Key metrics
- Context window
- 1,048,576 tokens
- Input price
- $0.30 / 1M tokens
- Output price
- $2.50 / 1M tokens
- Cached input
- $0.030 / 1M tokens
Official link:https://ai.google.dev/gemini-api/docs/pricing
Last verified: 2026-10-07
- · Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.
- · Gemini Developer API paid-tier text rates (not Vertex AI). Output includes thinking tokens. Cache storage is billed separately by token-hours; grounding and tools cost extra.
Multi-currency (per 1M tokens)
| Currency | Input | Cached | Output |
|---|---|---|---|
| USD | $0.30 | $0.03 | $2.50 |
| CNY | ¥2.15 | ¥0.21 | ¥17.88 |
| EUR | 0,28 € | 0,03 € | 2,30 € |
| JPY | ¥44 | ¥4 | ¥365 |
* Live search cost uses sources / 1000 × price and currently applies to xAI Grok only.
Frequently Asked Questions
What is the cost per 1M tokens for Google Gemini 3.5 Flash-Lite?
Google Gemini 3.5 Flash-Lite costs $0.30 per 1M input tokens and $2.50 per 1M output tokens, with cached input at $0.030 per 1M tokens.
How much does it cost per 1K tokens?
Per 1K tokens: $0.0003 for input and $0.0025 for output. This is useful for calculating costs for smaller workloads or individual API calls.
What is the estimated monthly cost for typical usage?
For 10M input + 2M output tokens per month across requests within the short-context tier, Google Gemini 3.5 Flash-Lite would cost approximately $8.00. Daily usage of 100K tokens (50K in, 50K out) costs about $0.1400.
Does Google Gemini 3.5 Flash-Lite offer a free tier?
Check Google's official documentation for free tier availability. Some providers offer free credits for new users or limited free usage. Visit https://ai.google.dev/gemini-api/docs/pricing for current free tier details.
How does prompt caching work to reduce costs?
With prompt caching enabled, input pricing drops to $0.030 per 1M tokens for repeated context (a 90% discount), while output remains $2.50 per 1M tokens. Caching is ideal for repeated prompts or system messages.
What is the context window size for Gemini 3.5 Flash-Lite?
Gemini 3.5 Flash-Lite supports a 1,048,576 token context window. This determines the maximum combined length of your input prompt and output response.
How frequently is this pricing information updated?
All prices reference official Google documentation (https://ai.google.dev/gemini-api/docs/pricing), last verified on 2026-10-07. Review the source before budgeting; other catalog entries may have older verification dates.
How can I calculate exact costs for my use case?
Use our free token calculator to estimate costs based on your specific usage pattern. The calculator supports all major models and shows costs in multiple currencies. You can also compare costs across different models to find the most economical option.