DeepSeek R1 0528 Qwen3 8B
ReasoningOpen weightsText8B parameters
DeepSeek R1 family model by DeepSeek.
About the DeepSeek R1 family
The open reasoning model that upended the industry in Jan 2025 — RL-trained chain-of-thought at ~$6M training cost, MIT-licensed, plus distilled versions from 1.5B to 70B that put reasoning on laptops. R1 1776 is Perplexity's decensored fine-tune.
All DeepSeek R1 modelsHow this one is configured
This configuration ~8B parameters.
