Overview
An 80B-total, 3B-active architecture preview combining hybrid linear attention and ultra-sparse MoE.
An 80B-total, 3B-active architecture preview combining hybrid linear attention and ultra-sparse MoE.
Only published specifications are shown.
An 80B-total, 3B-active architecture preview combining hybrid linear attention and ultra-sparse MoE.
Qwen3-Next-80B-A3B released open weights.
View sourceSelect a card to view the related model.