Skip to content
Yihua Zhang
Articles
Posts
EN
中
◐
Series
Anatomy of DeepSeek
Reading order
2025-01-20
From Zero to Reasoning Hero: How DeepSeek-R1 Leverages Reinforcement Learning to Master Complex Reasoning
Post-training
·
en / zh
·
2 equations
Anatomy of DeepSeek 1
GRPO
2025-02-02
Why Cache 32 Heads When One Latent Variable Suffices? A Theory-to-Code Guide to DeepSeek’s MLA for KV-Cache
Systems
·
en / zh
·
24 equations
Anatomy of DeepSeek 2
KV-Cache
2025-02-27
DualPipe Explained: A Comprehensive Guide to DualPipe That Anyone Can Understand—Even Without a Distributed Background
Systems
·
en / zh
Anatomy of DeepSeek 3
Distributed