Qwen3.8-Flash supports a default context window of 262,144 tokens, expandable to 1 million tokens, enabling it to handle large files, lengthy conversations and extensive research materials. It also released open-source weights for Qwen3.8-Flash-Next, allowing the developer community to evaluate an architecture that Qwen said would serve as a prototype for its next-generation Qwen4 model family.

Alibaba's Qwen team is teasing its next architecture a day early—and the specs say it runs near-frontier scale on a fraction of the power.

Alibaba’s Qwen team is scheduled to open-source Qwen3.8-Flash-Next and an FP8 version at 11 p.m. Beijing time on Aug. 26. The model is described as a