Gemini 3.5 Flash API

Call Gemini 3.5 Flash through Wokey and review published catalog pricing, upstream-native API forms, context, and model capabilities.

Model ID
gemini-3.5-flash
Vendors
google
Supply status
Current supply unavailable

Catalog pricing snapshot

Prices are per 1M tokens; request-time pricing controls settlement.

MeterWokeyOfficial referenceSavings
Input$0.3$1.580%
Output$1.8$980%
Cache read$0.03——
Cache write———

Official price source: https://ai.google.dev/gemini-api/docs/pricing

Model capabilities

Context
1,000,000 tokens
Max output
65,536 tokens
Streaming
Supported
Tools
Supported
Vision
Supported
Supported API forms
/v1/chat/completions

Send a request

/v1/chat/completions

curl https://api.wokey.ai/v1/chat/completions \
  -H "Authorization: Bearer $WOKEY_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"gemini-3.5-flash","messages":[{"role":"user","content":"Hello"}]}'

Use Gemini 3.5 Flash in your tools

The gateway is https://api.wokey.ai (OpenAI-compatible clients usually take https://api.wokey.ai/v1), and the model ID is gemini-3.5-flash. Open the guide for the client you use.

Frequently asked questions

How much does the Gemini 3.5 Flash API cost, and how much does it save versus the official API?

Wokey lists input at $0.3 per 1M tokens and output at $1.8 per 1M tokens; the official references are $1.5 and $9. That is 80% lower for input and 80% lower for output. Request-time pricing controls settlement, and the vendor pricing source is linked on this page.

What is the Gemini 3.5 Flash API model ID?

The model ID for Gemini 3.5 Flash on Wokey is gemini-3.5-flash. Use it exactly as written in the request body's model field or in your client's model setting (for example ANTHROPIC_MODEL in Claude Code, or model in the Codex CLI config.toml); GET https://api.wokey.ai/v1/models also lists it.

Which API forms does Gemini 3.5 Flash support, and how do I call it through Wokey?

Gemini 3.5 Flash supports /v1/chat/completions (Chat Completions requests). Set the API base URL to https://api.wokey.ai, authenticate with a Wokey API key, and use the canonical model ID gemini-3.5-flash; you can start with /v1/chat/completions. The Send a request section includes a curl example for every supported API form.

How do I use Gemini 3.5 Flash in opencode, Hermes Agent, and similar clients?

These clients call Gemini 3.5 Flash over OpenAI Chat Completions: point the API address at https://api.wokey.ai (clients that expect an OpenAI SDK-style address take https://api.wokey.ai/v1), add your Wokey API key, and use the model ID gemini-3.5-flash. Each client's guide shows its config format.

What are the context window and maximum output for Gemini 3.5 Flash?

The catalog context window is 1,000,000 tokens and the maximum output is 65,536 tokens. These are separate limits, and a request must also satisfy the upstream constraints in effect when it is processed.

How are cache reads and cache writes priced for Gemini 3.5 Flash?

Cache read: Wokey lists $0.03 per 1M tokens; the vendor does not publish a separate official reference. Cache write: the vendor does not publish a separate meter, so Wokey also shows — as the public price.

How does calling Gemini 3.5 Flash through Wokey differ from OpenRouter?

Wokey lists Gemini 3.5 Flash at $0.3 input / $1.8 output per 1M tokens, against an official reference of $1.5 / $9. OpenRouter states it adds no markup on inference and charges its fee when you buy credits (5.5% by card). Wokey carries a curated set of models and has no per-provider routing parameters, so OpenRouter fits better if you need a wider catalog. The OpenRouter alternative page has the full comparison.

Does Gemini 3.5 Flash support streaming, tool calling, and image input?

Streaming is supported; tool calling is supported; image input is supported. These values come directly from the current runtime model catalog and are not inferred from the model name.

Related models and documentation

Get API key · gemini-3.5-flash