DeepSeek R1 Distill Qwen 1.5B
ReasoningOpen weightsText1.5B parameters
DeepSeek R1 family model by DeepSeek.
About the DeepSeek R1 family
The open reasoning model that upended the industry in Jan 2025 — RL-trained chain-of-thought at ~$6M training cost, MIT-licensed, plus distilled versions from 1.5B to 70B that put reasoning on laptops. R1 1776 is Perplexity's decensored fine-tune.
All DeepSeek R1 modelsHow this one is configured
This configuration ~1.5B parameters.
