GPT-6 Luna API

Call GPT-6 Luna through Wokey and review dated catalog pricing, upstream-native API forms, context, and model capabilities.

Model ID
gpt-6-luna
Vendors
openai
Supply status
Currently callable

Catalog pricing snapshot

Prices are per 1M tokens; request-time pricing controls settlement.

MeterWokeyOfficial referenceSavings
Input$0.09$0.110%
Output$0.45$0.510%
Cache read$0.009$0.0110%
Cache write$0.1125$0.12510%

Price checked:

Official price source: https://developers.openai.com/api/docs/models/gpt-6-luna

Model capabilities

Context
1,050,000 tokens
Max output
128,000 tokens
Streaming
Supported
Tools
Supported
Vision
Supported
Supported API forms
/v1/chat/completions · /v1/responses

Send a request

/v1/chat/completions

curl https://api.wokey.ai/v1/chat/completions \
  -H "Authorization: Bearer $WOKEY_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"gpt-6-luna","messages":[{"role":"user","content":"Hello"}]}'

/v1/responses

curl https://api.wokey.ai/v1/responses \
  -H "Authorization: Bearer $WOKEY_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"gpt-6-luna","input":"Hello"}'

Frequently asked questions

How much does the GPT-6 Luna API cost, and how much does it save versus the official API?

As of 2026-09-23, Wokey lists input at $0.09 per 1M tokens and output at $0.45 per 1M tokens; the official references are $0.1 and $0.5. That is 10% lower for input and 10% lower for output. Request-time pricing controls settlement, and the vendor pricing source is linked on this page.

Which API forms does GPT-6 Luna support, and how do I call it through Wokey?

GPT-6 Luna supports /v1/chat/completions (Chat Completions requests) and /v1/responses (Responses requests). Set the API base URL to https://api.wokey.ai, authenticate with a Wokey API key, and use the canonical model ID gpt-6-luna; you can start with /v1/responses. The Send a request section includes a curl example for every supported API form.

What are the context window and maximum output for GPT-6 Luna?

The catalog context window is 1,050,000 tokens and the maximum output is 128,000 tokens. These are separate limits, and a request must also satisfy the upstream constraints in effect when it is processed.

How are cache reads and cache writes priced for GPT-6 Luna?

Cache read: Wokey lists $0.009 per 1M tokens and the official reference is $0.01 per 1M tokens. Cache write: Wokey lists $0.1125 per 1M tokens and the official reference is $0.125 per 1M tokens.

How do GPT-6 Luna and GPT-6 Astra differ in price and context?

In this price snapshot, GPT-6 Luna lists input/output at $0.09 / $0.45 with a 1,050,000-token context window; GPT-6 Astra lists $0.9 / $4.5 with a 1,050,000-token context window. This is a factual price-and-capacity comparison, not a quality ranking; also compare their supported API forms and capabilities.

Related models and documentation

Get API key · gpt-6-luna