StudyArena

DeepSeek R1 0528 Qwen3 8B

ReasoningOpen weightsText8B parameters

DeepSeek R1 family model by DeepSeek.

About the DeepSeek R1 family

The open reasoning model that upended the industry in Jan 2025 — RL-trained chain-of-thought at ~$6M training cost, MIT-licensed, plus distilled versions from 1.5B to 70B that put reasoning on laptops. R1 1776 is Perplexity's decensored fine-tune.

All DeepSeek R1 models

How this one is configured

This configuration ~8B parameters.

Who makes it

DeepSeekChinaHangzhou lab famous for frontier-level open models at rock-bottom cost.

More from DeepSeek R1

← Back to all models