DevTk.AI

GPT-6 Luna

OpenAI

Updated October 2026. GPT-6 Luna by OpenAI: $0.1/M cache-miss input, $0.5/M output tokens. Cached input: $0.01/M. Explicit cache writes: $0.125/M. Long-context pricing at or above 272K input tokens: $0.2/M input, $0.75/M output. 1.1M context, 128K max output. Vision & Function Calling. Free calculator + compare 40+ models.

Input Price

$0.1

cache miss / 1M tokens

Cached Input

$0.01

per 1M tokens

Output Price

$0.5

per 1M tokens

Context Window

1.1M

tokens

Specifications

ProviderOpenAI
Model IDgpt-6-luna
Input Price$0.1 / 1M cache-miss tokens
Cached Input Price$0.01 / 1M tokens
Explicit Cache Write Price$0.125 / 1M tokens
Output Price$0.5 / 1M tokens
Long Context Threshold272K input tokens
Long Context Pricing$0.2 / 1M cache-miss input, $0.02 / 1M cached input, $0.25 / 1M explicit cache writes, $0.75 / 1M output
Context Window1.1M tokens
Max Output128K tokens
Capabilities
textvisionfunction_callingstructured_output
Release Date2026-09-22
Pricing SourceOfficial OpenAI pricing
Price Verified2026-10-02 · GPT-6.1 Sol is $2/$10 with $0.10/M cached input (half GPT-6 Sol). GPT-6 Astra also offers Ultrafast at 6x Standard. GPT-5.6 Sol promotional $4/$20 pricing lasts at least through 2026-11-21.
NotesGPT-6 model for focused, repeatable, high-volume work. Prompts above 272K input tokens bill the entire request at long-context rates; Batch and Flex cost 50% of Standard, while Fast mode costs 2x.

Monthly Cost Estimates

Estimated monthly costs based on different daily usage levels (assuming 50% input / 50% output split). Input estimates use cache-miss pricing, so cache-heavy workloads can be lower.

Daily TokensMonthly CostAnnual Cost
10K $0.09 $1.08
50K $0.45 $5.40
100K $0.9 $10.80
500K $4.50 $54.00
1.0M $9.00 $108.00

About GPT-6 Luna

GPT-6 Luna is a large language model by OpenAI. It features a 1.1M token context window with up to 128K tokens of output per request. The model supports 4 capabilities: text, vision, function_calling, structured_output.

At $0.1 per million cache-miss input tokens and $0.5 per million output tokens, GPT-6 Luna is positioned as a cost-effective option in the OpenAI lineup. Repeated prefix input can be charged at $0.01 per million cached tokens. Use our Token Counter to estimate how many tokens your prompts use, and our Pricing Calculator to compare costs across all models.

GPT-6 Luna Key Details

  • Pricing: $0.1/M cache-miss input tokens, $0.01/M cached input tokens, $0.125/M explicit cache writes, $0.5/M output tokens
  • Context window: 1.1M tokens — one of the largest available
  • Max output: 128K tokens per response
  • Capabilities: text, vision, function_calling, structured_output
  • Highlights: GPT-6 model for focused, repeatable, high-volume work. Prompts above 272K input tokens bill the entire request at long-context rates; Batch and Flex cost 50% of Standard, while Fast mode costs 2x.
  • Released: 2026-09-22

Other OpenAI Models

Similar Price Range

Related Tools

FAQ

How much does GPT-6 Luna cost?

GPT-6 Luna costs $0.1 per million cache-miss input tokens and $0.5 per million output tokens. Cached input costs $0.01 per million tokens. For a typical workload of 100K input tokens/day and 50K output tokens/day, expect approximately $1.05/month before cache-hit savings.

What is GPT-6 Luna's context window?

GPT-6 Luna supports a context window of 1.1M tokens. This means your combined input prompt and output response can be up to 1.1M tokens. The maximum output per response is 128K tokens.

Is GPT-6 Luna good for my use case?

GPT-6 Luna supports text, vision, function_calling, structured_output. As a budget-friendly model, it works well for high-volume tasks like classification, summarization, and simple generation. Use our Pricing Calculator to compare with alternatives.