Luchuan Song
发表论文 10 篇 · 总被引 78 次 · h-index 2
代表论文
- Video Understanding With Large Language Models: A Survey (2025 · IEEE Transactions on Circuits and Systems for Video Technology · 被引 70)
- Video Understanding with Large Language Models: A Survey (2023 · arXiv (Cornell University) · 被引 8)
- Omni-Judge: Can Omni-LLMs Serve as Human-Aligned Judges for Text-Conditioned Audio-Video Generation? (2026 · arXiv (Cornell University))
- Omni-Judge: Can Omni-LLMs Serve as Human-Aligned Judges for Text-Conditioned Audio-Video Generation? (2026 · Open MIND)
- Video-LMM Post-Training: A Deep Dive into Video Reasoning with Large Multimodal Models (2025 · arXiv (Cornell University))
- MMPerspective: Do MLLMs Understand Perspective? A Comprehensive Benchmark for Perspective Perception, Reasoning, and Robustness (2025 · arXiv (Cornell University))