Renrui Zhang
发表论文 43 篇 · 总被引 3635 次 · h-index 20
代表论文
- Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis (2024 · Computer Vision and Pattern Recognition · 被引 1494)
- MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI (2024 · International Conference on Machine Learning · 被引 207)
- T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT (2025 · Neural Information Processing Systems · 被引 165)
- Can We Generate Images with CoT? Let's Verify and Reinforce Image Generation Step by Step (2025 · arXiv.org · 被引 138)
- MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency (2025 · International Conference on Machine Learning · 被引 131)
- CoMat: Aligning Text-to-Image Diffusion Model with Image-to-Text Concept Matching (2024 · Neural Information Processing Systems · 被引 73)