All changelog entries
FeatureNew0.13.2ModelsVision

Qwen3.8-Flash — a third vision model, and the cheapest of them

August 27, 2026

Qwen3.8-Flash is live. It reads text, images and video, carries a 1M-token context window, and is the least expensive of the three vision models we added this week.

Alibaba positions it for coding assistance, agentic workflows, document and codebase analysis, chart reading and long-video analysis. Call it as Qwen3.8-Flash.

It costs 1.5 requests of quota per call and is available on every paid plan and the free tier — the same terms as GLM-5.3-Flash and DeepSeek-V4-Flash-Vision.

Alibaba has not published benchmark scores for this model yet. We will add them to the model page once they exist rather than borrowing figures from its siblings.