DeepSeek-R1-Zero
ModelReasoning model trained through large-scale reinforcement learning without supervised fine-tuning.
- Jan 20, 2025
- —
- —
- Yes
- —
- 1
Releases
Jan 20, 2025DeepSeek-R1-Zero weights releasedOpen weights
DeepSeek published its pure reinforcement-learning reasoning checkpoint.

