StudyArena

DeepSeek R1

9 models · DeepSeek

The open reasoning model that upended the industry in Jan 2025 — RL-trained chain-of-thought at ~$6M training cost, MIT-licensed, plus distilled versions from 1.5B to 70B that put reasoning on laptops. R1 1776 is Perplexity's decensored fine-tune.

Models in this family

← Back to all models