DeepSeek-V4.1-Flash API
Call DeepSeek-V4.1-Flash through Wokey and review dated catalog pricing, upstream-native API forms, context, and model capabilities.
- Model ID
deepseek-flash- Vendors
- deepseek, opencode
- Supply status
- Currently callable
Catalog pricing snapshot
Prices are per 1M tokens; request-time pricing controls settlement.
| Meter | Wokey | Official reference | Savings |
|---|---|---|---|
| Input | $0.112 | $0.14 | 20% |
| Output | $0.448 | $0.56 | 20% |
| Cache read | $0.00224 | $0.0028 | 20% |
| Cache write | — | — | — |
Price checked:
Official price source: https://api-docs.deepseek.com/quick_start/pricing
Model capabilities
- Context
- 1,000,000 tokens
- Max output
- 393,216 tokens
- Streaming
- Supported
- Tools
- Supported
- Vision
- Supported
- Supported API forms
/v1/chat/completions·/v1/responses·/v1/messages
Send a request
/v1/chat/completions
curl https://api.wokey.ai/v1/chat/completions \
-H "Authorization: Bearer $WOKEY_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"deepseek-flash","messages":[{"role":"user","content":"Hello"}]}'/v1/responses
curl https://api.wokey.ai/v1/responses \
-H "Authorization: Bearer $WOKEY_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"deepseek-flash","input":"Hello"}'/v1/messages
curl https://api.wokey.ai/v1/messages \
-H "x-api-key: $WOKEY_API_KEY" \
-H "anthropic-version: 2023-06-01" \
-H "Content-Type: application/json" \
-d '{"model":"deepseek-flash","max_tokens":1024,"messages":[{"role":"user","content":"Hello"}]}'Frequently asked questions
How much does the DeepSeek-V4.1-Flash API cost, and how much does it save versus the official API?
As of 2026-09-14, Wokey lists input at $0.112 per 1M tokens and output at $0.448 per 1M tokens; the official references are $0.14 and $0.56. That is 20% lower for input and 20% lower for output. Request-time pricing controls settlement, and the vendor pricing source is linked on this page.
Which API forms does DeepSeek-V4.1-Flash support, and how do I call it through Wokey?
DeepSeek-V4.1-Flash supports /v1/chat/completions (Chat Completions requests), /v1/responses (Responses requests), and /v1/messages (Messages requests). Set the API base URL to https://api.wokey.ai, authenticate with a Wokey API key, and use the canonical model ID deepseek-flash; you can start with /v1/chat/completions. The Send a request section includes a curl example for every supported API form.
What are the context window and maximum output for DeepSeek-V4.1-Flash?
The catalog context window is 1,000,000 tokens and the maximum output is 393,216 tokens. These are separate limits, and a request must also satisfy the upstream constraints in effect when it is processed.
How are cache reads and cache writes priced for DeepSeek-V4.1-Flash?
Cache read: Wokey lists $0.00224 per 1M tokens and the official reference is $0.0028 per 1M tokens. Cache write: the vendor does not publish a separate meter, so Wokey also shows — as the public price.
How do DeepSeek-V4.1-Flash and DeepSeek V4 Flash differ in price and context?
In this price snapshot, DeepSeek-V4.1-Flash lists input/output at $0.112 / $0.448 with a 1,000,000-token context window; DeepSeek V4 Flash lists $0.112 / $0.448 with a 1,000,000-token context window. This is a factual price-and-capacity comparison, not a quality ranking; also compare their supported API forms and capabilities.
Related models and documentation
Get API key · deepseek-flash