Overview
Used large-scale reinforcement learning to shape reasoning and released both full and distilled model weights.
Used large-scale reinforcement learning to shape reasoning and released both full and distilled model weights.
Only published specifications are shown.
Used large-scale reinforcement learning to shape reasoning and released both full and distilled model weights.
R1 brought open-weight reasoning models into mainstream global comparison and accelerated replication of RL-based reasoning recipes.
DeepSeek-R1 and its distilled models were released with open weights.
View sourceSelect a card to view the related model.