Shulin Tian
发表论文 17 篇 · 总被引 295 次 · h-index 6
代表论文
- AHA: A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation (2024 · International Conference on Learning Representations · 被引 140)
- Ego-R1: Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning (2025 · arXiv.org · 被引 56)
- MMInA: Benchmarking Multihop Multimodal Internet Agents (2024 · Annual Meeting of the Association for Computational Linguistics · 被引 50)
- Evaluation Agent: Efficient and Promptable Evaluation Framework for Visual Generative Models (2024 · Annual Meeting of the Association for Computational Linguistics · 被引 40)
- Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark (2025 · Annual Meeting of the Association for Computational Linguistics · 被引 20)
- A Simple Baseline for Streaming Video Understanding (2026 · arXiv.org · 被引 12)