Qwen 3.8 Flash via the Tokenator API

What the model is for

Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.

Specification

Full nameQwen 3.8 Flash
API IDqwen-3.8-flash
Model vendorQwen
TypeText model (chat / completions)
Context window1M tokens
Max output131.1K tokens
Image inputyes
Tokenator usage multiplier1.7×
Input modalitiestext, image, video
Aliasescursor-qwen-3.8-flash
Supported API formats/v1/chat/completions /v1/responses /v1/messages
Upstream providers1
Current statusavailable
Data updated2026-08-27

Upstream providers

ProviderMultiplier
Unified LLM API1.7×

A request goes to the first available provider by priority; if it fails, Tokenator switches to the next one.

Example request

curl
curl https://api.tokenator.top/v1/chat/completions \
  -H "Authorization: Bearer sk-your-tokenator-key" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen-3.8-flash",
    "messages": [{"role": "user", "content": "Hello"}]
  }'

Compatible tools

The model can be used in any tool that accepts a custom base URL: Claude Code, Codex CLI, OpenCode, Cursor, Cline, Kilo Code, Cherry Studio, OpenClaw.

FAQ about Qwen 3.8 Flash

What is the API ID of Qwen 3.8 Flash?

qwen-3.8-flash — put this into the model field of your request.

What context window does Qwen 3.8 Flash have?

1M tokens; max output is 131.1K tokens.

How many tokens will a request cost?

(input + output) × 1.7×. See the documentation for details.

Which tools support this model?

Any tool that accepts a custom base URL: Claude Code, Codex CLI, OpenCode, Cursor, Cline, Kilo Code. Setup guides are in the integrations section.