Z.ai's new GLM-5.3-Flash model offers strong value for the money, comes with clear weaknesses, and brings a noteworthy infrastructure angle.
GLM-5.3-Flash is the first natively multimodal model in the GLM-5 series, according to Z.ai. It has 320 billion total parameters, of which only 18 billion are active, ships under an MIT license, and offers a context window of one million tokens. The weights are available on Hugging Face.
Measurements from Artificial Analysis put the model at 57 points on the Intelligence Index at maximum reasoning effort. That's just three points behind the larger GLM-5.3, which scores 60, and level with GPT-5.6 Terra and Muse Spark 1.2.
The price is what stands out. Cost per task on the index runs 0.09 dollars, against 0.68 dollars for GLM-5.3, roughly 7.5 times cheaper. That puts the model on the Pareto frontier of intelligence and cost, according to Artificial Analysis, and adds it to a growing list of Chinese models that have recently put heavy price pressure on Western providers.
On the Intelligence Index, GLM-5.3-Flash lands at 57 points and sits in the most attractive cost-versus-intelligence quadrant. | Image: Artificial Analysis













