DeepSeek R1
9 models · DeepSeek
The open reasoning model that upended the industry in Jan 2025 — RL-trained chain-of-thought at ~$6M training cost, MIT-licensed, plus distilled versions from 1.5B to 70B that put reasoning on laptops. R1 1776 is Perplexity's decensored fine-tune.
Models in this family
- DeepSeek R1 (Jan '25)DeepSeekOpenReasoning
- DeepSeek R1 0528 (May '25)DeepSeekOpenReasoning
- DeepSeek R1 0528 Qwen3 8BDeepSeekOpen8B parametersReasoning
- DeepSeek R1 Distill Llama 70BDeepSeekOpen70B parametersReasoning
- DeepSeek R1 Distill Llama 8BDeepSeekOpen8B parametersReasoning
- DeepSeek R1 Distill Qwen 1.5BDeepSeekOpen1.5B parametersReasoning
- DeepSeek R1 Distill Qwen 14BDeepSeekOpen14B parametersReasoning
- DeepSeek R1 Distill Qwen 32BDeepSeekOpen32B parametersReasoning
- R1 1776DeepSeekOpenReasoning
