Toggle navigation

Z.ai

Available in Shortly AI

GLM 5.3 Flash

Released August 26, 2026

ReasoningVision (Image)Tool Use

GLM-5.3-Flash is Z.ai’s native multimodal coding model, featuring 320B total parameters, 18B activated parameters, and a 1M-token context window. Its efficient hybrid attention architecture supports visual coding, tool use, and end-to-end professional workflows across code, browsers, documents, and graphical interfaces.

Context window
1M tokens
Maximum output
131K tokens
Input pricing
$0.15/M
Output pricing
$0.50/M

Independent benchmark

LiveBench performance

71.6overall

Benchmark data from LiveBench, release 2026-06-25. Scores are pinned and are not fetched on page load.

GLM 5.3 Flash LiveBench category scores
CategoryScore
Reasoning77.644
Coding78.950
Agentic Coding56.768
Mathematics81.245
Data Analysis76.401
Language77.307
Instruction Following52.821

Benchmarks measure selected evaluation tasks and do not guarantee performance for every prompt or workflow.

FAQ

Similar AI models

Available with one subscription

Start with GLM 5.3 Flash in Shortly AI

Keep your conversations, files and model choices together in one workspace.