Yao Lu
发表论文 18 篇 · 总被引 2175 次 · h-index 15
代表论文
- VILA: On Pre-training for Visual Language Models (2023 · Computer Vision and Pattern Recognition · 被引 908)
- LongVILA: Scaling Long-Context Visual Language Models for Long Videos (2024 · International Conference on Learning Representations · 被引 314)
- VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation (2024 · International Conference on Learning Representations · 被引 283)
- NVILA: Efficient Frontier Visual Language Models (2024 · Computer Vision and Pattern Recognition · 被引 229)
- EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos (2025 · arXiv.org · 被引 122)
- RADIOv2.5: Improved Baselines for Agglomerative Vision Foundation Models (2024 · Computer Vision and Pattern Recognition · 被引 91)