Live drop
GLM-5.3-Flash
Z.aiDropped Aug 26, 2026
Natively multimodal 320B-A18B MoE with hybrid sparse/linear attention for low-cost agentic coding and vision tasks.
About this release
Live drop
Z.aiDropped Aug 26, 2026
Natively multimodal 320B-A18B MoE with hybrid sparse/linear attention for low-cost agentic coding and vision tasks.
About this release
GLM-5.3-Flash is Z.ai's first natively multimodal model in the GLM-5 series, released under MIT license with open weights. It uses a hybrid sparse + linear attention architecture (320B total, 18B active) plus Manifold-Constrained Hyper-Connections on a 30T multimodal pretraining corpus, achieving 57 on the Artificial Analysis Intelligence Index while approaching Claude Opus 4.8 on coding and agentic benchmarks at a fraction of the cost of GLM-5.3, with 1M context and native image/video support.
Read the official Z.ai launch pageFirst builds incoming — this lane fills in automatically as people start building with GLM-5.3-Flash. Know one? Submit the URL.