Z.ai releases GLM-5.3-Flash, a 320B-A18B natively multimodal MoE with 1M context, MIT weights, and $0.15/M input pricing.

Z.ai's GLM-5.3-Flash runs natively on Chinese AI chips, hits 63.4 on DeepSWE, and offers API access at $0.15 per million input tokens.

For the past week, developers have been puzzling over a model called Ox Alpha. It appeared on...