Verified 2026-10-07 · sourced from Google
Gemini 3.8 Flash Token Calculator, Pricing & 100K/1M Cost
Check Google Gemini 3.8 Flash pricing, estimate 100K and 1M token cost, and size a real API budget before you send a single request. Standard pricing is $0.75 per million input tokens and $3.75 per million output tokens with a 1M token context window.
Quick answer: Gemini 3.8 Flash pricing per 1M tokens is $0.75 input and $3.75 output. Context window: 1,048,576 tokens · Cached input: $0.075 / 1M.
Best for searches like Gemini 3.8 Flash token calculator, Gemini 3.8 Flash pricing, Gemini 3.8 Flash 100K tokens cost, Gemini 3.8 Flash 1M token cost.
Reference rates: Promotional through 2026-12-31 · Developer API
USD per million text tokens. Verified 2026-10-07 · Official source
| Mode / prompt size | Input | Cache read | Cache write | Output |
|---|---|---|---|---|
| Standard | $0.75 | $0.075 | Not estimated | $3.75 |
| Batch | $0.375 | $0.0375 | Not estimated | $1.875 |
Cache storage: $0.5 per million tokens per hour, charged separately and excluded from the per-call estimate.
1,048,576 context tokens · max output 65,536. Local token counts are estimates; use the provider’s usage counts for an invoice estimate.
Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.
Gemini Developer API paid-tier text rates (not Vertex AI). Output includes thinking tokens. Cache storage is billed separately by token-hours; grounding and tools cost extra.
Promotion: standard $0.75 input / $0.075 cache read / $3.75 output through 2026-12-31. From 2027-01-01: $1.50 / $0.15 / $7.50. Batch $0.375 / $0.0375 / $1.875 becomes $0.75 / $0.075 / $3.75. Storage rises from $0.50 to $1 per million token-hours. Calculator switches at 00:00 UTC on 2027-01-01.
Pick the route that matches what you searched for
Some visitors want a fast Gemini 3.8 Flash API cost estimate, others want a direct 100K or 1M token budget, and some are already comparing alternatives. These shortcuts remove the extra click.
Estimate a single request or prompt budget right now.
Jump straight to the most common budgeting checkpoint.
Use this when you are sizing production traffic or a monthly plan.
Open the closest head-to-head comparison instead of researching from scratch.
Context window
1,048,576 tokens
Input price
$0.75 / 1M
Output price
$3.75 / 1M
Cached input
$0.075 / 1M
Pricing modes and thresholds
Usage scenarios
Compare standard and cached pricing (where available) across common workloads.
| Scenario | Tokens in | Tokens out | Total tokens | Standard cost | Cached cost |
|---|---|---|---|---|---|
Quick chat reply Single user question with a short assistant answer | 650 | 220 | 870 | $0.0013 | $0.0009 |
Coding assistant session Multi-turn pair programming exchange (≈6 turns) | 2,600 | 1,400 | 4,000 | $0.0072 | $0.0054 |
Knowledge base response Retrieval-augmented answer referencing multiple passages | 12,000 | 3,000 | 15,000 | $0.0203 | $0.0121 |
Near-max context run Large document processing approaching the 1M token limit | 983,040 | 65,536 | 1,048,576 | $0.983 | $0.319 |
Daily & monthly budgeting
Translate usage into predictable operating expenses across popular deployment sizes.
| Profile | Requests/day | Tokens/day | Daily cost | Monthly cost | Cached daily | Cached monthly |
|---|---|---|---|---|---|---|
| Team pilot | 25 | 75,000 | $0.131 | $3.94 | $0.0975 | $2.93 |
| Product launch | 100 | 500,000 | $0.825 | $24.75 | $0.589 | $17.66 |
| Enterprise scale | 500 | 3,000,000 | $5.25 | $157.50 | $3.90 | $117.00 |
Pricing notes
- Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.
- Gemini Developer API paid-tier text rates (not Vertex AI). Output includes thinking tokens. Cache storage is billed separately by token-hours; grounding and tools cost extra.
- Promotion: standard $0.75 input / $0.075 cache read / $3.75 output through 2026-12-31. From 2027-01-01: $1.50 / $0.15 / $7.50. Batch $0.375 / $0.0375 / $1.875 becomes $0.75 / $0.075 / $3.75. Storage rises from $0.50 to $1 per million token-hours. Calculator switches at 00:00 UTC on 2027-01-01.
Frequently asked questions
How much does Gemini 3.8 Flash cost per 1,000 tokens?
At the published rates of $0.75 per million input tokens and $3.75 per million output tokens, a typical 1,000 token request (≈70% input, 30% output) costs about $0.0016.
Does Gemini 3.8 Flash offer cached input discounts?
Gemini 3.8 Flash drops input costs to $0.075 per million cached tokens. Using cached contexts, that same 1,000 token call totals $0.0012, a significant saving for chatbots and RAG systems.
What is the context window for Gemini 3.8 Flash?
Gemini 3.8 Flash supports up to 1,048,576 tokens (1M), allowing large prompts and retrieval-augmented payloads in a single call.
How fresh is the Gemini 3.8 Flash pricing data?
Pricing is sourced from https://ai.google.dev/gemini-api/docs/pricing and was last verified on 2026-10-07. The calculator updates automatically when models.json is refreshed.