Z.ai released and open-sourced GLM-5.3-Flash, a natively multimodal mixture-of-experts GLM model with a 1M-token context
Open source release Provisional 72% confidence first seen
Z.ai announced the release of GLM-5.3-Flash, describing it as a natively multimodal mixture-of-experts model with a very large (1,048,576-token) context window. Coverage also states that the model was open-sourced and shared alongside performance and deployment notes, including pricing details and claims about using domestically produced AI chips during early testing.