R. Xu
发表论文 4 篇 · 总被引 15752 次 · h-index 4
代表论文
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models (2024 · arXiv.org · 被引 8191)
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning (2025 · Nature · 被引 5514)
- DeepSeek-V3 Technical Report (2024 · arXiv.org · 被引 3874)
- Let the Expert Stick to His Last: Expert-Specialized Fine-Tuning for Sparse Architectural Large Language Models (2024 · Conference on Empirical Methods in Natural Language Processing · 被引 21)