All changelog entries
FeatureNew0.13.1ModelsVision

Two new flash models — and both of them can see

August 26, 2026

GLM-5.3-Flash and DeepSeek-V4-Flash-Vision are live. Both take images as well as text, and both carry a 1M-token context window.

GLM-5.3-Flash is built for coding and long-running agent work — it scores 84.3 on Terminal-Bench 2.1, within a point of the frontier models, and also accepts video input. DeepSeek-V4-Flash-Vision adds image understanding to the V4 Flash line, which posts 79% on SWE-bench Verified and 88.4% on LiveCodeBench, and is aimed at documents, charts and multimodal agent workflows.

Both are available on every paid plan and on the free tier, and cost 1.5 requests of quota per call. Call them as GLM-5.3-Flash and DeepSeek-V4-Flash-Vision.

Benchmark figures are as published by the model authors; DeepSeek's vision variant is marked experimental upstream.