Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Ganqu Cui

发表论文 82 篇 · 总被引 13687 次 · h-index 34

代表论文

  • The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models (2025 · arXiv.org · 被引 402)
  • TTRL: Test-Time Reinforcement Learning (2025 · Neural Information Processing Systems · 被引 217)
  • A Survey of Reinforcement Learning for Large Reasoning Models (2025 · arXiv.org · 被引 160)
  • A Survey of Efficient Reasoning for Large Reasoning Models: Language, Multimodality, and Beyond (2025 · arXiv.org · 被引 142)
  • MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe (2025 · arXiv.org · 被引 127)
  • Intern-S1: A Scientific Multimodal Foundation Model (2025 · arXiv.org · 被引 81)