Smallest, fastest GPT-6 model for high-volume classification, extraction and chat. Long-context requests bill at $0.20 input / $0.75 output per 1M. Built by OpenAI.
Prices updated
Input price
$0.10
per 1M tokens · Standard
Output price
$0.50
per 1M tokens · Standard
Input limit
1.1M
tokens
Output limit
128K
tokens
Input formats
Output formats
GPT-6 Luna price history
2 price records since 9/25/2026
GPT-6 Luna cost calculator
$4.50/month
Standard pricing
Overview
What is GPT-6 Luna?
GPT-6 Luna is a text & reasoning and code & slms and vision model from OpenAI. Smallest, fastest GPT-6 model for high-volume classification, extraction and chat. Long-context requests bill at $0.20 input / $0.75 output per 1M. Its 1.1M context window and $0.10 input price make it a candidate for cost-sensitive, high-throughput applications.
Benchmarks
GPT-6 Luna benchmarks & speed
How GPT-6 Luna scores on standardized evaluations, and where it lands among every model we track.
38.1
Intelligence Index
141tok/s
Output speed
1277
DesignArena Elo
Reasoning
HLE
38.5%
Humanity's Last Exam · top 25%
Coding
SciCode
54.6%
Python for scientific computing · top 29%
Latency & design
- Time to first token
- 95.97s
- DesignArena win rate
- 48.5%
- Design battles judged
- 4,539
Independent scores from Artificial Analysis and DesignArena · updated 10/8/2026. Higher is better; ranks compare against every model we track with that score.
Rates
GPT-6 Luna pricing
Pricing is based on token usage. Provider-specific caching, batch, regional, and tool charges may affect the final cost.
Input · Standard
$0.10
Per 1M tokens
Cached input · Standard
$0.01
Per 1M tokens
Capabilities
GPT-6 Luna Tools
Tools available when using GPT-6 Luna through supported provider APIs.
Function calling
Supported
Structured outputs
Supported
JSON mode
Supported
Reasoning
Supported
Built-in web search
Not supported
Log probabilities
Not supported
Deterministic seed
Supported
Parallel tool calls
Not supported
Prompt caching
Supported
Where to run it
GPT-6 Luna API providers
7 providers serve GPT-6 Luna. Prices are per 1M tokens; uptime is the last 24 hours.
| Provider | Input | Output | Cached | Context | Uptime |
|---|---|---|---|---|---|
| OpenAI | $0.05 | $0.25 | $0.0050 | 1.1M | 98.61% |
| OpenAI | $0.10 | $0.50 | $0.01 | 1.1M | 99.95% |
| Azure | $0.10 | $0.50 | $0.01 | 1.1M | 99.92% |
| Azure | $0.11 | $0.55 | $0.01 | 1.1M | 99.99% |
| Azure | $0.11 | $0.55 | $0.01 | 1.1M | 99.92% |
| Amazon Bedrock | $0.11 | $0.55 | $0.01 | 1.1M | 98.59% |
| OpenAI | $0.20 | $1.00 | $0.02 | 1.1M | 99.93% |
- Input
- $0.05
- Output
- $0.25
- Cached
- $0.0050
- Context
- 1.1M
- Uptime
- 98.61%
- Input
- $0.10
- Output
- $0.50
- Cached
- $0.01
- Context
- 1.1M
- Uptime
- 99.95%
- Input
- $0.10
- Output
- $0.50
- Cached
- $0.01
- Context
- 1.1M
- Uptime
- 99.92%
- Input
- $0.11
- Output
- $0.55
- Cached
- $0.01
- Context
- 1.1M
- Uptime
- 99.99%
- Input
- $0.11
- Output
- $0.55
- Cached
- $0.01
- Context
- 1.1M
- Uptime
- 99.92%
- Input
- $0.11
- Output
- $0.55
- Cached
- $0.01
- Context
- 1.1M
- Uptime
- 98.59%
- Input
- $0.20
- Output
- $1.00
- Cached
- $0.02
- Context
- 1.1M
- Uptime
- 99.93%
6 of 7
Strengths and limitations
Strengths
- Low input cost at $0.10 per million tokens suits high-volume workloads.
- 1.1M context supports large documents, repositories and extended conversations.
- Supports text & reasoning and code & slms and vision workloads in one model.
Limitations
- Generated tokens cost 5× more than input tokens, which matters for verbose responses.
- Large context capacity does not guarantee consistent retrieval across the entire prompt.
- Arena Elo and MMLU-Pro are directional; test accuracy, latency and reliability on your own workload before committing.
Arena Elo, MMLU-Pro and pricing figures are illustrative; validate current vendor terms before purchase.
Same provider
