Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Tao Gui

发表论文 206 篇 · 总被引 9047 次 · h-index 42

代表论文

  • AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning (2025 · arXiv.org · 被引 67)
  • BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping (2025 · arXiv.org · 被引 51)
  • CL-bench: A Benchmark for Context Learning (2026 · arXiv.org · 被引 45)
  • Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models (2026 · Computer Science Review · 被引 20)
  • OctoBench: Benchmarking Scaffold-Aware Instruction Following in Repository-Grounded Agentic Coding (2026 · Annual Meeting of the Association for Computational Linguistics · 被引 16)
  • MagicGUI: A Foundational Mobile GUI Agent with Scalable Data Pipeline and Reinforcement Fine-tuning (2025 · arXiv.org · 被引 16)