While Alibaba has kept inference and token price low, enterprises need to consider other metrics to determine if this is the right model for them.

August 26, 2026

Chinese AI tech giant Alibaba introduced a new variant of its flagship model on Wednesday, aiming to compete on price with U.S. vendors and to provide enterprises with lower-cost AI models. However, the latest Alibaba release serves as a reminder to enterprises that price is not always the main driver of model choice.

Qwen 3.8-Flash-Next is a multimodal model that provides an early preview of the architecture used in the upcoming Qwen 4. It is a 125B-parameter model, with an active 6B parameters per token and a mixture-of-experts (MoE) architecture. In comparison, Qwen 3.8 Max has 2.4 trillion parameters. Qwen 3.8-27B has 27B parameters. Rival Chinese AI vendor Moonshot’s Kimi K3 MoE model has 2.8 trillion parameters.

Alibaba highlighted that compared to Qwen 3.7-Plus, the Flash-Next version has lower training and inference costs. The model excels at computer use and can interact with complex APIs, calculators and custom external databases using visual and text prompts, according to the vendor