Z.ai releases GLM-5.3-Flash with native visual capabilities and hybrid architecture
Original titleGLM-5.3-Flash
AISummary
Z.ai has released GLM-5.3-Flash, a model with native visual capabilities that observe interfaces, rendering results, and interaction feedback across code, browsers, and GUIs.
It uses a hybrid linear and sparse attention architecture with 320B total parameters and 18B activated, which the company says significantly reduces compute and KV-cache requirements.
The release notes also describe support for office document and financial research workflows.
AIWhy it matters
The release notes give GLM-5.3-Flash's architecture, parameter counts, and cybersecurity findings, which make the model's scope concrete for comparison with earlier GLM releases.
Source: Z.ai Release Notes · docs.z.aiPublished · added here