Overview
The first natively multimodal GLM-5 Flash model: 320B total, 18B active parameters with hybrid sparse and linear attention.
The first natively multimodal GLM-5 Flash model: 320B total, 18B active parameters with hybrid sparse and linear attention.
Only published specifications are shown.
The first natively multimodal GLM-5 Flash model: 320B total, 18B active parameters with hybrid sparse and linear attention.
It uses a new base model and 1M context to deliver GLM-5 long-horizon agent capability at lower inference cost.
GLM-5.3-Flash launched with MIT-licensed open weights.
View sourceSelect a card to view the related model.