Skip to content
AImpact
IT EN
← Years

2026

86 entries

High

Alibaba releases Qwen 3.5: sparse multimodal MoE, 262K context, Apache 2.0

On 16 February 2026 Alibaba opened Qwen 3.5, starting with the 397B-A17B flagship: sparse MoE with 512 experts of which 11 are active, Gated DeltaNet attention, multimodal input, and 262,144 tokens of native context (around 1M with YaRN). Apache 2.0 and 201 languages; first among open-weight models on instruction following and multilinguality, still behind on reasoning.

Open Source Models QwenOpen WeightsMultilingual