Zhen-fei Yin
发表论文 37 篇 · 总被引 2332 次 · h-index 19
代表论文
- WorldSimBench: Towards Video Generation Models as World Simulators (2024 · arXiv.org · 被引 1251)
- LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark (2023 · Neural Information Processing Systems · 被引 235)
- MP5: A Multi-modal Open-ended Embodied System in Minecraft via Active Perception (2023 · Computer Vision and Pattern Recognition · 被引 96)
- SPA-VL: A Comprehensive Safety Preference Alignment Dataset for Vision Language Models (2024 · Computer Vision and Pattern Recognition · 被引 87)
- MineDreamer: Learning to Follow Instructions via Chain-of-Imagination for Simulated-World Control (2024 · arXiv.org · 被引 58)
- Are We There Yet? Revealing the Risks of Utilizing Large Language Models in Scholarly Peer Review (2024 · arXiv.org · 被引 57)