Verified 2026-10-07 · sourced from Google

Gemini 3.8 Flash Token Calculator, Pricing & 100K/1M Cost

Check Google Gemini 3.8 Flash pricing, estimate 100K and 1M token cost, and size a real API budget before you send a single request. Standard pricing is $0.75 per million input tokens and $3.75 per million output tokens with a 1M token context window.

Quick answer: Gemini 3.8 Flash pricing per 1M tokens is $0.75 input and $3.75 output. Context window: 1,048,576 tokens · Cached input: $0.075 / 1M.

Best for searches like Gemini 3.8 Flash token calculator, Gemini 3.8 Flash pricing, Gemini 3.8 Flash 100K tokens cost, Gemini 3.8 Flash 1M token cost.

Reference rates: Promotional through 2026-12-31 · Developer API

USD per million text tokens. Verified 2026-10-07 · Official source

Mode / prompt sizeInputCache readCache writeOutput
Standard$0.75$0.075Not estimated$3.75
Batch$0.375$0.0375Not estimated$1.875

Cache storage: $0.5 per million tokens per hour, charged separately and excluded from the per-call estimate.

1,048,576 context tokens · max output 65,536. Local token counts are estimates; use the provider’s usage counts for an invoice estimate.

Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.

Gemini Developer API paid-tier text rates (not Vertex AI). Output includes thinking tokens. Cache storage is billed separately by token-hours; grounding and tools cost extra.

Promotion: standard $0.75 input / $0.075 cache read / $3.75 output through 2026-12-31. From 2027-01-01: $1.50 / $0.15 / $7.50. Batch $0.375 / $0.0375 / $1.875 becomes $0.75 / $0.075 / $3.75. Storage rises from $0.50 to $1 per million token-hours. Calculator switches at 00:00 UTC on 2027-01-01.

Pick the route that matches what you searched for

Some visitors want a fast Gemini 3.8 Flash API cost estimate, others want a direct 100K or 1M token budget, and some are already comparing alternatives. These shortcuts remove the extra click.

Context window

1,048,576 tokens

Input price

$0.75 / 1M

Output price

$3.75 / 1M

Cached input

$0.075 / 1M

Pricing modes and thresholds

Standard
$0.7500 input · $0.0750 cached · $3.7500 output
Batch
$0.3750 input · $0.0375 cached · $1.8750 output

Usage scenarios

Compare standard and cached pricing (where available) across common workloads.

ScenarioTokens inTokens outTotal tokensStandard costCached cost
Quick chat reply
Single user question with a short assistant answer
650220870$0.0013$0.0009
Coding assistant session
Multi-turn pair programming exchange (≈6 turns)
2,6001,4004,000$0.0072$0.0054
Knowledge base response
Retrieval-augmented answer referencing multiple passages
12,0003,00015,000$0.0203$0.0121
Near-max context run
Large document processing approaching the 1M token limit
983,04065,5361,048,576$0.983$0.319

Daily & monthly budgeting

Translate usage into predictable operating expenses across popular deployment sizes.

ProfileRequests/dayTokens/dayDaily costMonthly costCached dailyCached monthly
Team pilot2575,000$0.131$3.94$0.0975$2.93
Product launch100500,000$0.825$24.75$0.589$17.66
Enterprise scale5003,000,000$5.25$157.50$3.90$117.00

Pricing notes

  • Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.
  • Gemini Developer API paid-tier text rates (not Vertex AI). Output includes thinking tokens. Cache storage is billed separately by token-hours; grounding and tools cost extra.
  • Promotion: standard $0.75 input / $0.075 cache read / $3.75 output through 2026-12-31. From 2027-01-01: $1.50 / $0.15 / $7.50. Batch $0.375 / $0.0375 / $1.875 becomes $0.75 / $0.075 / $3.75. Storage rises from $0.50 to $1 per million token-hours. Calculator switches at 00:00 UTC on 2027-01-01.

Frequently asked questions

How much does Gemini 3.8 Flash cost per 1,000 tokens?

At the published rates of $0.75 per million input tokens and $3.75 per million output tokens, a typical 1,000 token request (≈70% input, 30% output) costs about $0.0016.

Does Gemini 3.8 Flash offer cached input discounts?

Gemini 3.8 Flash drops input costs to $0.075 per million cached tokens. Using cached contexts, that same 1,000 token call totals $0.0012, a significant saving for chatbots and RAG systems.

What is the context window for Gemini 3.8 Flash?

Gemini 3.8 Flash supports up to 1,048,576 tokens (1M), allowing large prompts and retrieval-augmented payloads in a single call.

How fresh is the Gemini 3.8 Flash pricing data?

Pricing is sourced from https://ai.google.dev/gemini-api/docs/pricing and was last verified on 2026-10-07. The calculator updates automatically when models.json is refreshed.

Related resources

Other Token Calculators

Explore More Tools