Zhihong Shao
发表论文 27 篇 · 总被引 21017 次 · h-index 15
代表论文
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models (2024 · arXiv.org · 被引 8191)
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning (2025 · Nature · 被引 5514)
- DeepSeek-V3 Technical Report (2024 · arXiv.org · 被引 3874)
- DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model (2024 · arXiv.org · 被引 1376)
- Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations (2023 · Annual Meeting of the Association for Computational Linguistics · 被引 1090)
- CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing (2023 · International Conference on Learning Representations · 被引 852)