DeepSeek-AI published research on DeepSeek-R1-Zero and DeepSeek-R1, alongside work on distilling reasoning capabilities into smaller models.
Why it mattered: the paper contributed methods and results for studying how training choices affect multi-step reasoning.
Date note: this entry marks the paper’s first arXiv submission, not the earlier model announcement.
Book a consultation