Hongxu Yin
发表论文 72 篇 · 总被引 8302 次 · h-index 40
代表论文
- SpatialRGPT: Grounded Spatial Reasoning in Vision Language Model (2024 · Neural Information Processing Systems · 被引 368)
- LongVILA: Scaling Long-Context Visual Language Models for Long Videos (2024 · International Conference on Learning Representations · 被引 314)
- NaVILA: Legged Robot Vision-Language-Action Model for Navigation (2024 · Robotics · 被引 243)
- NVILA: Efficient Frontier Visual Language Models (2024 · Computer Vision and Pattern Recognition · 被引 229)
- EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos (2025 · arXiv.org · 被引 122)
- GDPO: Group reward-Decoupled Normalization Policy Optimization for Multi-reward RL Optimization (2026 · arXiv.org · 被引 108)