Gemini 3.5 Flash via the Tokenator API
What the model is for
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed.
Specification
| Full name | Gemini 3.5 Flash |
| API ID | gemini-3.5-flash |
| Model vendor | Gemini |
| Type | Text model (chat / completions) |
| Context window | 1M tokens |
| Max output | 65.5K tokens |
| Image input | yes |
| Tokenator usage multiplier | 1.7× – 2.1× |
| Input modalities | text, image, audio, video, file |
| Aliases | cursor-gemini-3.5-flash |
| Supported API formats | /v1/chat/completions /v1/responses /v1/messages |
| Upstream providers | 3 |
| Current status | available |
| Data updated | 2026-08-04 |
Benchmarks
| Intelligence index | 50.2 |
| Coding index | 70.1 |
| Agentic index | 37.4 |
Artificial Analysis data loaded into the catalog for this model.
Upstream providers
| Provider | Multiplier |
|---|---|
| Gemini Partner #1 | 1.7× |
| Gemini Partner #3 | 2.1× |
| AiPartners Ltd | 1.7× |
A request goes to the first available provider by priority; if it fails, Tokenator switches to the next one.
Example request
curl https://api.tokenator.top/v1/chat/completions \ -H "Authorization: Bearer sk-your-tokenator-key" \ -H "Content-Type: application/json" \ -d '{ "model": "gemini-3.5-flash", "messages": [{"role": "user", "content": "Hello"}] }'
Compatible tools
The model can be used in any tool that accepts a custom base URL: Claude Code, Codex CLI, OpenCode, Cursor, Cline, Kilo Code, OpenClaw.
See also
FAQ about Gemini 3.5 Flash
What is the API ID of Gemini 3.5 Flash?
gemini-3.5-flash — put this into the model field of your request.
What context window does Gemini 3.5 Flash have?
1M tokens; max output is 65.5K tokens.
How many tokens will a request cost?
(input + output) × 1.7× – 2.1×. See the documentation for details.
Which tools support this model?
Any tool that accepts a custom base URL: Claude Code, Codex CLI, OpenCode, Cursor, Cline, Kilo Code. Setup guides are in the integrations section.