Verified 2026-10-07 · sourced from OpenAI

GPT-6 Luna Token Calculator, Pricing & 100K/1M Cost

Check OpenAI GPT-6 Luna pricing, estimate 100K and 1M token cost, and size a real API budget before you send a single request. Standard pricing is $0.10 per million input tokens and $0.50 per million output tokens with a 1.1M token context window.

Quick answer: GPT-6 Luna pricing per 1M tokens is $0.10 input and $0.50 output. Context window: 1,050,000 tokens · Cached input: $0.010 / 1M.

Best for searches like GPT-6 Luna token calculator, GPT-6 Luna pricing, GPT-6 Luna 100K tokens cost, GPT-6 Luna 1M token cost.

Reference rates: Standard · ≤272K input · global

USD per million text tokens. Verified 2026-10-07 · Official source

Mode / prompt sizeInputCache readCache writeOutput
Standard ≤272,000 input$0.1$0.01$0.125$0.5
Standard >272,000 input$0.2$0.02$0.25$0.75
Batch ≤272,000 input$0.05$0.005$0.0625$0.25
Batch >272,000 input$0.1$0.01$0.125$0.375
Flex ≤272,000 input$0.05$0.005$0.0625$0.25
Flex >272,000 input$0.1$0.01$0.125$0.375
Fast mode ≤272,000 input$0.2$0.02$0.25$1
Fast mode >272,000 input$0.4$0.04$0.5$1.5

1,050,000 context tokens · max output 128,000. Local token counts are estimates; use the provider’s usage counts for an invoice estimate.

Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.

Above 272,000 input tokens, the full request uses 2× input/cache rates and 1.5× output rates. Batch/Flex and Fast/Ultrafast are separate processing options, not stackable discounts.

Global endpoint rates shown. Regional processing adds 10% where available and is excluded from this estimate.

Billed output includes reasoning/thinking tokens that cannot be inferred from pasted final text; use API-reported output usage for accurate billing.

Pick the route that matches what you searched for

Some visitors want a fast GPT-6 Luna API cost estimate, others want a direct 100K or 1M token budget, and some are already comparing alternatives. These shortcuts remove the extra click.

Context window

1,050,000 tokens

Input price

$0.10 / 1M

Output price

$0.50 / 1M

Cached input

$0.010 / 1M

Pricing modes and thresholds

Standard
$0.1000 input · $0.0100 cached · $0.5000 output
Batch
$0.0500 input · $0.0050 cached · $0.2500 output
Flex
$0.0500 input · $0.0050 cached · $0.2500 output
Fast mode
$0.2000 input · $0.0200 cached · $1.0000 output

Long-context pricing starts above 272,000 input tokens.

Usage scenarios

Compare standard and cached pricing (where available) across common workloads.

ScenarioTokens inTokens outTotal tokensStandard costCached cost
Quick chat reply
Single user question with a short assistant answer
650220870$0.0002$0.0001
Coding assistant session
Multi-turn pair programming exchange (≈6 turns)
2,6001,4004,000$0.0010$0.0007
Knowledge base response
Retrieval-augmented answer referencing multiple passages
12,0003,00015,000$0.0027$0.0016
Near-max context run
Large document processing approaching the 1.1M token limit
924,000126,0001,050,000$0.279$0.113

Daily & monthly budgeting

Translate usage into predictable operating expenses across popular deployment sizes.

ProfileRequests/dayTokens/dayDaily costMonthly costCached dailyCached monthly
Team pilot2575,000$0.0175$0.525$0.0130$0.390
Product launch100500,000$0.182$5.47$0.119$3.58
Enterprise scale5003,000,000$1.15$34.50$0.790$23.70

Pricing notes

  • Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.
  • Above 272,000 input tokens, the full request uses 2× input/cache rates and 1.5× output rates. Batch/Flex and Fast/Ultrafast are separate processing options, not stackable discounts.
  • Global endpoint rates shown. Regional processing adds 10% where available and is excluded from this estimate.
  • Billed output includes reasoning/thinking tokens that cannot be inferred from pasted final text; use API-reported output usage for accurate billing.

Frequently asked questions

How much does GPT-6 Luna cost per 1,000 tokens?

At the published rates of $0.10 per million input tokens and $0.50 per million output tokens, a typical 1,000 token request (≈70% input, 30% output) costs about $0.0002.

Does GPT-6 Luna offer cached input discounts?

GPT-6 Luna drops input costs to $0.010 per million cached tokens. Using cached contexts, that same 1,000 token call totals $0.0002, a significant saving for chatbots and RAG systems.

What is the context window for GPT-6 Luna?

GPT-6 Luna supports up to 1,050,000 tokens (1.1M), allowing large prompts and retrieval-augmented payloads in a single call.

How fresh is the GPT-6 Luna pricing data?

Pricing is sourced from https://developers.openai.com/api/docs/models/gpt-6-luna and was last verified on 2026-10-07. The calculator updates automatically when models.json is refreshed.

Related resources

Other Token Calculators

Explore More Tools