Qwen3.8 Flash

qwen3.8-flash · Qwen

Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding, office tasks, long-context reasoning, and agent workflows. It supports a 1M-token context, 128K output, web access, and tool calling. Compared with Qwen3.7-Plus, Qwen3.8-Flash significantly reduces training and inference costs—the training overhead is only about one-ninth of the former—while offering stronger capabilities on coding and office tasks.

API Pricing

Input$0.1126 / 1M tokens
Output$0.38 / 1M tokens
Cache read$0.0141 / 1M tokens

Specifications

Context1M tokens
Max output131K tokens
Modalitiestext, image, video
CapabilitiesThinking, Streaming, Tool calling, Web search, Code interpreter, Structured outputs, Prompt caching

Frequently asked questions

What is Qwen3.8 Flash?

Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding, office tasks, long-context reasoning, and agent workflows. It supports a 1M-token context, 128K output, web access, and tool calling. Compared with Qwen3.7-Plus, Qwen3.8-Flash significantly reduces training and inference costs—the training overhead is only about one-ninth of the former—while offering stronger capabilities on coding and office tasks.

What is the context length of Qwen3.8 Flash?

Qwen3.8 Flash has a 1,000,000 token context window. It supports up to 131,072 output tokens.

How much does Qwen3.8 Flash cost?

On AIHubMix, Qwen3.8 Flash costs $0.1126 per million input tokens and $0.38 per million output tokens. Cached input reads are billed at $0.0141 per million tokens.

What modalities does Qwen3.8 Flash support?

Qwen3.8 Flash accepts text, image and video input.

What capabilities does Qwen3.8 Flash support?

Qwen3.8 Flash supports tool calling, function calling, structured outputs, web search, long context and thinking. Per-protocol parameter support is listed in the capability table on this page.

How do I call Qwen3.8 Flash via API?

Qwen3.8 Flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3.8-flash — no other code changes needed.

Who created Qwen3.8 Flash?

Qwen3.8 Flash is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was Qwen3.8 Flash released?

Qwen3.8 Flash was released on August 26, 2026 by Qwen.

More models from Qwen

See all Qwen models →

Qwen3.8 Max 2026 09-02

by Qwen

Qwen3.8-Max-0902 (also known as qwen3.8-max-2026-09-02) is a snapshot version of Alibaba…

$1.69/1M in · $5.07/1M out
991,000 tokens context

Wan3.0 Video

by Qwen

Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model…

$2/1M in · $2/1M out

Wan3.0 Video Prime

by Qwen

Wan3.0 Video Prime is Alibaba Cloud’s preview high-speed edition of its All-in-One video…

$2/1M in · $2/1M out

Qwen3.8 Max

by Qwen

Qwen3.8-Max is Alibaba Cloud Tongyi Qianwen's next-generation flagship large language…

$1.69/1M in · $5.07/1M out
991,000 tokens context

Qwen3.8 2.4t A95B

by Qwen

Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a…

$2/1M in · $6/1M out
262,000 tokens context

Qwen Image 3.0

by Qwen

Qwen Image 3.0(qwen-image-3.0) is an image generation and editing model developed by…

$2/1M in

Use Qwen3.8 Flash via the AIHubMix unified API — one interface for every major LLM.