Grok 4.6 API

Call Grok 4.6 through Wokey and review dated catalog pricing, upstream-native API forms, context, and model capabilities.

Model ID
grok-4.6
Vendors
xai
Supply status
Currently callable

Catalog pricing snapshot

Prices are per 1M tokens; request-time pricing controls settlement.

MeterWokeyOfficial referenceSavings
Input$0.4$280%
Output$1.2$680%
Cache read$0.1$0.580%
Cache write

Price checked:

Official price source: https://docs.x.ai/developers/models/grok-4.6

Model capabilities

Context
500,000 tokens
Max output
500,000 tokens
Streaming
Supported
Tools
Supported
Vision
Supported
Supported API forms
/v1/chat/completions · /v1/responses

Send a request

/v1/chat/completions

curl https://api.wokey.ai/v1/chat/completions \
  -H "Authorization: Bearer $WOKEY_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"grok-4.6","messages":[{"role":"user","content":"Hello"}]}'

/v1/responses

curl https://api.wokey.ai/v1/responses \
+  -H "Authorization: Bearer $WOKEY_API_KEY" \
+  -H "Content-Type: application/json" \
+  -d '{"model":"grok-4.6","input":"Hello"}'

Frequently asked questions

How much does the Grok 4.6 API cost, and how much does it save versus the official API?

As of 2026-08-26, Wokey lists input at $0.4 per 1M tokens and output at $1.2 per 1M tokens; the official references are $2 and $6. That is 80% lower for input and 80% lower for output. Request-time pricing controls settlement, and the vendor pricing source is linked on this page.

Which API forms does Grok 4.6 support, and how do I call it through Wokey?

Grok 4.6 supports /v1/chat/completions (Chat Completions requests) and /v1/responses (Responses requests). Set the API base URL to https://api.wokey.ai, authenticate with a Wokey API key, and use the canonical model ID grok-4.6; you can start with /v1/chat/completions. The Send a request section includes a curl example for every supported API form.

What are the context window and maximum output for Grok 4.6?

The catalog context window is 500,000 tokens and the maximum output is 500,000 tokens. These are separate limits, and a request must also satisfy the upstream constraints in effect when it is processed.

How are cache reads and cache writes priced for Grok 4.6?

Cache read: Wokey lists $0.1 per 1M tokens and the official reference is $0.5 per 1M tokens. Cache write: the vendor does not publish a separate meter, so Wokey also shows — as the public price.

How do Grok 4.6 and Grok 4.5 differ in price and context?

In this price snapshot, Grok 4.6 lists input/output at $0.4 / $1.2 with a 500,000-token context window; Grok 4.5 lists $0.4 / $1.2 with a 500,000-token context window. This is a factual price-and-capacity comparison, not a quality ranking; also compare their supported API forms and capabilities.

Related models and documentation

Get API key · grok-4.6