Gemini 3.5 Flash API
Call Gemini 3.5 Flash through Wokey and review published catalog pricing, upstream-native API forms, context, and model capabilities.
- Model ID
gemini-3.5-flash- Vendors
- Supply status
- Current supply unavailable
Catalog pricing snapshot
Prices are per 1M tokens; request-time pricing controls settlement.
| Meter | Wokey | Official reference | Savings |
|---|---|---|---|
| Input | $0.3 | $1.5 | 80% |
| Output | $1.8 | $9 | 80% |
| Cache read | $0.03 | — | — |
| Cache write | — | — | — |
Official price source: https://ai.google.dev/gemini-api/docs/pricing
Model capabilities
- Context
- 1,000,000 tokens
- Max output
- 65,536 tokens
- Streaming
- Supported
- Tools
- Supported
- Vision
- Supported
- Supported API forms
/v1/chat/completions
Send a request
/v1/chat/completions
curl https://api.wokey.ai/v1/chat/completions \
-H "Authorization: Bearer $WOKEY_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"gemini-3.5-flash","messages":[{"role":"user","content":"Hello"}]}'Use Gemini 3.5 Flash in your tools
The gateway is https://api.wokey.ai (OpenAI-compatible clients usually take https://api.wokey.ai/v1), and the model ID is gemini-3.5-flash. Open the guide for the client you use.
- Use Gemini 3.5 Flash in opencode: OpenAI Chat Completions · base URL and API key setup
- Use Gemini 3.5 Flash in OpenClaw: OpenAI Chat Completions · base URL and API key setup
- Use Gemini 3.5 Flash in Hermes Agent: OpenAI Chat Completions · base URL and API key setup
- OpenRouter alternative: Wokey vs OpenRouter: Compare token prices, payment fees, API forms, and upstream sources.
- Verifiable AI API: which models and API forms carry response proofs: Response proofs currently cover only raw passthrough Claude Messages and GPT Responses on official routes.
Frequently asked questions
How much does the Gemini 3.5 Flash API cost, and how much does it save versus the official API?
Wokey lists input at $0.3 per 1M tokens and output at $1.8 per 1M tokens; the official references are $1.5 and $9. That is 80% lower for input and 80% lower for output. Request-time pricing controls settlement, and the vendor pricing source is linked on this page.
What is the Gemini 3.5 Flash API model ID?
The model ID for Gemini 3.5 Flash on Wokey is gemini-3.5-flash. Use it exactly as written in the request body's model field or in your client's model setting (for example ANTHROPIC_MODEL in Claude Code, or model in the Codex CLI config.toml); GET https://api.wokey.ai/v1/models also lists it.
Which API forms does Gemini 3.5 Flash support, and how do I call it through Wokey?
Gemini 3.5 Flash supports /v1/chat/completions (Chat Completions requests). Set the API base URL to https://api.wokey.ai, authenticate with a Wokey API key, and use the canonical model ID gemini-3.5-flash; you can start with /v1/chat/completions. The Send a request section includes a curl example for every supported API form.
How do I use Gemini 3.5 Flash in opencode, Hermes Agent, and similar clients?
These clients call Gemini 3.5 Flash over OpenAI Chat Completions: point the API address at https://api.wokey.ai (clients that expect an OpenAI SDK-style address take https://api.wokey.ai/v1), add your Wokey API key, and use the model ID gemini-3.5-flash. Each client's guide shows its config format.
What are the context window and maximum output for Gemini 3.5 Flash?
The catalog context window is 1,000,000 tokens and the maximum output is 65,536 tokens. These are separate limits, and a request must also satisfy the upstream constraints in effect when it is processed.
How are cache reads and cache writes priced for Gemini 3.5 Flash?
Cache read: Wokey lists $0.03 per 1M tokens; the vendor does not publish a separate official reference. Cache write: the vendor does not publish a separate meter, so Wokey also shows — as the public price.
How does calling Gemini 3.5 Flash through Wokey differ from OpenRouter?
Wokey lists Gemini 3.5 Flash at $0.3 input / $1.8 output per 1M tokens, against an official reference of $1.5 / $9. OpenRouter states it adds no markup on inference and charges its fee when you buy credits (5.5% by card). Wokey carries a curated set of models and has no per-provider routing parameters, so OpenRouter fits better if you need a wider catalog. The OpenRouter alternative page has the full comparison.
Does Gemini 3.5 Flash support streaming, tool calling, and image input?
Streaming is supported; tool calling is supported; image input is supported. These values come directly from the current runtime model catalog and are not inferred from the model name.
Related models and documentation
Get API key · gemini-3.5-flash