StudyArena

DeepSeek R1 Distill Llama 70B

ReasoningOpen weightsText70B parameters

DeepSeek R1 family model by DeepSeek.

About the DeepSeek R1 family

The open reasoning model that upended the industry in Jan 2025 — RL-trained chain-of-thought at ~$6M training cost, MIT-licensed, plus distilled versions from 1.5B to 70B that put reasoning on laptops. R1 1776 is Perplexity's decensored fine-tune.

All DeepSeek R1 models

How this one is configured

This configuration ~70B parameters.

Who makes it

DeepSeekChinaHangzhou lab famous for frontier-level open models at rock-bottom cost.

More from DeepSeek R1

← Back to all models