qwen3.5-flash · Qwen
The Qwen3.5 native vision-language Flash series models are designed with a hybrid architecture that integrates linear attention mechanisms and sparse mixture-of-experts models, achieving higher inference efficiency. Compared with the 3 series, the models deliver leapfrog improvements in both pure-text and multimodal performance; they respond quickly and combine inference speed with high performance.
The Qwen3.5 native vision-language Flash series models are designed with a hybrid architecture that integrates linear attention mechanisms and sparse mixture-of-experts models, achieving higher inference efficiency. Compared with the 3 series, the models deliver leapfrog improvements in both pure-text and multimodal performance; they respond quickly and combine inference speed with high performance.
Qwen3.5 Flash has a 1,000,000 token context window. It supports up to 65,536 output tokens.
On AIHubMix, Qwen3.5 Flash costs $0.0282 per million input tokens and $0.282 per million output tokens. Cached input reads are billed at $0.0028 per million tokens.
Qwen3.5 Flash accepts text, image and video input.
Qwen3.5 Flash supports tool calling, function calling, structured outputs, web search, long context and thinking. Per-protocol parameter support is listed in the capability table on this page.
Qwen3.5 Flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to qwen3.5-flash — no other code changes needed.
Qwen3.5 Flash is developed by Qwen. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Qwen3.8-Max-0902 (also known as qwen3.8-max-2026-09-02) is a snapshot version of Alibaba…
Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding…
Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model…
Wan3.0 Video Prime is Alibaba Cloud’s preview high-speed edition of its All-in-One video…
Qwen3.8-Max is Alibaba Cloud Tongyi Qianwen's next-generation flagship large language…
Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a…
Use Qwen3.5 Flash via the AIHubMix unified API — one interface for every major LLM.