06:01
2026-09-18
zenmux.ai
large-language-models
New Model Available: GLM 5.3 FlashX
Z.ai released GLM-5.3-FlashX, a native multimodal model that delivers inference speeds of up to 200 tokens/s, faster than GLM-5.3-Flash. The model uses a hybrid sparse and linear attention architectur…