Overview
An experimental model built on V3.1-Terminus, introducing DeepSeek Sparse Attention for long-context efficiency.
An experimental model built on V3.1-Terminus, introducing DeepSeek Sparse Attention for long-context efficiency.
Only published specifications are shown.
An experimental model built on V3.1-Terminus, introducing DeepSeek Sparse Attention for long-context efficiency.
DeepSeek-V3.2-Exp released model weights.
View sourceSelect a card to view the related model.