Overview
The high-capability DeepSeek V4 preview: a 1.6T-total, 49B-active MoE with a 1M context window.
The high-capability DeepSeek V4 preview: a 1.6T-total, 49B-active MoE with a 1M context window.
Only published specifications are shown.
The high-capability DeepSeek V4 preview: a 1.6T-total, 49B-active MoE with a 1M context window.
It introduced V4’s sparse-attention, million-token architecture and released the weights alongside the preview.
DeepSeek released V4-Pro Preview with open weights.
View sourceSelect a card to view the related model.