Alibaba released Qwen3.8 Flash, a new multimodal 125B-parameter MoE model. 

> 262K native context, extensible to 1M with YaRN.
> Pricing on QwenCloud API - &#036; 0.16/1M input tokens and &#036; 0.47/1M output tokens.
> Built on top of the new architecture, serving as a precursor to the architecture used in Qwen4.
> Qwen3.8 Flash scores 58.7 on DeepSWE 1.1 and 62.5 on SWE-bench Pro.