Model index
DeepSeek·DeepSeek V / R

DeepSeek-R1

Used large-scale reinforcement learning to shape reasoning and released both full and distilled model weights.

CurrentReasoningCodingOpen weights
MODEL NUMBER018
Release date
2025-01-20
Organization
DeepSeek
Family
DeepSeek V / R
01

Specifications

Only published specifications are shown.

Total parameters
671B
Active parameters
37B
Architecture
MoE
Context
128K
Input modalities
Text
Open weights
Yes
License
MIT
Status
Current
02

Overview

Overview

Used large-scale reinforcement learning to shape reasoning and released both full and distilled model weights.

Why it mattered

R1 brought open-weight reasoning models into mainstream global comparison and accelerated replication of RL-based reasoning recipes.

03

Changes from predecessor

Compared with DeepSeek-V3

PreviousDeepSeek-V3671B · 128K ctx
CurrentDeepSeek-R1671B · 128K ctx
New capabilities Reasoning Chain of thought
04

Architecture & capabilities

Attention

MLA

Key capabilities

ReasoningCodingChain of thought
05

Release history

Open weights

DeepSeek-R1 and its distilled models were released with open weights.

View source
06

Related models

Select a card to view the related model.

07

Sources