DeepSeek-AI published research on DeepSeek-R1-Zero and DeepSeek-R1, alongside work on distilling reasoning capabilities into smaller models.

Why it mattered: the paper contributed methods and results for studying how training choices affect multi-step reasoning.

Date note: this entry marks the paper’s first arXiv submission, not the earlier model announcement.