Yuxiao Dong
发表论文 94 篇 · 总被引 10598 次 · h-index 41
代表论文
- GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning (2025 · 被引 354)
- GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning (2025 · arXiv.org · 被引 209)
- TreeRL: LLM Reinforcement Learning with On-Policy Tree Search (2025 · Annual Meeting of the Association for Computational Linguistics · 被引 74)
- MotionBench: Benchmarking and Improving Fine-Grained Video Motion Understanding for Vision Language Models (2025 · Computer Vision and Pattern Recognition · 被引 73)
- AgentRL: Scaling Agentic Reinforcement Learning with a Multi-Turn, Multi-Task Framework (2025 · arXiv.org · 被引 48)
- DeepDive: Advancing Deep Search Agents with Knowledge Graphs and Multi-Turn RL (2025 · arXiv.org · 被引 47)