DeepSeek V4 Flash API

Call DeepSeek V4 Flash through Wokey and review published catalog pricing, upstream-native API forms, context, and model capabilities.

Model ID
deepseek-v4-flash
Vendors
volcengine, deepseek, opencode
Supply status
Currently callable

Catalog pricing snapshot

Prices are per 1M tokens; request-time pricing controls settlement.

MeterWokeyOfficial referenceSavings
Input$0.112$0.1420%
Output$0.448$0.5620%
Cache read$0.00224$0.002820%
Cache write———

Official price source: https://api-docs.deepseek.com/quick_start/pricing

Model capabilities

Context
1,000,000 tokens
Max output
384,000 tokens
Streaming
Supported
Tools
Supported
Vision
Not supported
Supported API forms
/v1/chat/completions

Send a request

/v1/chat/completions

curl https://api.wokey.ai/v1/chat/completions \
  -H "Authorization: Bearer $WOKEY_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"deepseek-v4-flash","messages":[{"role":"user","content":"Hello"}]}'

Use DeepSeek V4 Flash in your tools

The gateway is https://api.wokey.ai (OpenAI-compatible clients usually take https://api.wokey.ai/v1), and the model ID is deepseek-v4-flash. Open the guide for the client you use.

Frequently asked questions

How much does the DeepSeek V4 Flash API cost, and how much does it save versus the official API?

Wokey lists input at $0.112 per 1M tokens and output at $0.448 per 1M tokens; the official references are $0.14 and $0.56. That is 20% lower for input and 20% lower for output. Request-time pricing controls settlement, and the vendor pricing source is linked on this page.

What is the DeepSeek V4 Flash API model ID?

The model ID for DeepSeek V4 Flash on Wokey is deepseek-v4-flash. Use it exactly as written in the request body's model field or in your client's model setting (for example ANTHROPIC_MODEL in Claude Code, or model in the Codex CLI config.toml); GET https://api.wokey.ai/v1/models also lists it.

Which API forms does DeepSeek V4 Flash support, and how do I call it through Wokey?

DeepSeek V4 Flash supports /v1/chat/completions (Chat Completions requests). Set the API base URL to https://api.wokey.ai, authenticate with a Wokey API key, and use the canonical model ID deepseek-v4-flash; you can start with /v1/chat/completions. The Send a request section includes a curl example for every supported API form.

How do I use DeepSeek V4 Flash in opencode, Hermes Agent, and similar clients?

These clients call DeepSeek V4 Flash over OpenAI Chat Completions: point the API address at https://api.wokey.ai (clients that expect an OpenAI SDK-style address take https://api.wokey.ai/v1), add your Wokey API key, and use the model ID deepseek-v4-flash. Each client's guide shows its config format.

What are the context window and maximum output for DeepSeek V4 Flash?

The catalog context window is 1,000,000 tokens and the maximum output is 384,000 tokens. These are separate limits, and a request must also satisfy the upstream constraints in effect when it is processed.

How are cache reads and cache writes priced for DeepSeek V4 Flash?

Cache read: Wokey lists $0.00224 per 1M tokens and the official reference is $0.0028 per 1M tokens. Cache write: the vendor does not publish a separate meter, so Wokey also shows — as the public price.

How does calling DeepSeek V4 Flash through Wokey differ from OpenRouter?

Wokey lists DeepSeek V4 Flash at $0.112 input / $0.448 output per 1M tokens, against an official reference of $0.14 / $0.56. OpenRouter states it adds no markup on inference and charges its fee when you buy credits (5.5% by card). Wokey carries a curated set of models and has no per-provider routing parameters, so OpenRouter fits better if you need a wider catalog. The OpenRouter alternative page has the full comparison.

How do DeepSeek V4 Flash and DeepSeek-V4.1-Flash differ in price and context?

In the published price comparison, DeepSeek V4 Flash lists input/output at $0.112 / $0.448 with a 1,000,000-token context window; DeepSeek-V4.1-Flash lists $0.112 / $0.448 with a 1,000,000-token context window. This is a factual price-and-capacity comparison, not a quality ranking; also compare their supported API forms and capabilities.

Related models and documentation

Get API key · deepseek-v4-flash