Skip to content
AIMarketCap

DeepSeek-R1-Zero

Model
DeepSeekFamily DeepSeek R1active

Reasoning model trained through large-scale reinforcement learning without supervised fine-tuning.

Released
Jan 20, 2025
Context window
—
Modality
—
Open source
Yes
API available
—
Releases
1

Releases

1
Jan 20, 2025DeepSeek-R1-Zero weights releasedDeepSeek-R1-Zero

DeepSeek published its pure reinforcement-learning reasoning checkpoint.