Higher benchmark scores don't mean lower cost. Qwen 3.8-Max and Claude Opus 5 both show it — and cost per successful task is the metric that catches it.

According to the company, Qwen3.8-Max can autonomously complete software projects lasting more than 10 days, reproduce research papers involving thousands of lines of code...

Alibaba has released Qwen3.8-Max, its largest and most capable AI model to date. The headline...