DeepSeek-R1: open-weights reasoning model rivals o1 and shakes markets
DeepSeek released R1 under the MIT license, a reasoning model matching OpenAI o1 on math and coding benchmarks, and showed with R1-Zero that reasoning can emerge from pure RL; on 27 January 2025 it topped the US App Store and NVIDIA lost ~$589B in market value in a single day.
Key facts
- Released 20 January 2025; paper arXiv 2501.12948
- MIT license, with distilled smaller models (1.5B–70B) based on Qwen and Llama
- R1-Zero trained with RL (GRPO) without supervised fine-tuning
- NVIDIA shares fell ~17% on 27 January 2025, erasing ~$589B — the largest one-day loss in US market history
- Peer-reviewed version published in Nature in September 2025
What happened
DeepSeek openly published a reasoning model and its RL recipe; its free chatbot app went viral worldwide.
Why it matters
The 'DeepSeek moment' showed that frontier reasoning could be replicated cheaply and openly, triggering a market shock, a wave of open reasoning models, and US policy debates on China.
Changelog
- 2026-09-29: created
Related events
- DeepSeek-V3: frontier-level open model trained for ~$5.6M in GPU time ★★★★★
- OpenAI o1: reasoning models trained with reinforcement learning ★★★★★
- Meta releases Llama 4 Scout and Maverick ★★★
- Moonshot AI releases Kimi K2, a 1-trillion-parameter open-weights agentic model ★★★
- OpenAI releases gpt-oss, its first open-weight LLMs since GPT-2 ★★★
Sources (3)
- paperDeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning (arXiv)
- codedeepseek-ai/DeepSeek-R1 (code & weights)
- pressNvidia sheds almost $600 billion in market cap (CNBC)
id: 2025-01-20-deepseek-r1 · updated 2026-09-29 · open in the interactive timeline