Sifei Liu
发表论文 25 篇 · 总被引 1248 次 · h-index 11
代表论文
- SpatialRGPT: Grounded Spatial Reasoning in Vision Language Model (2024 · Neural Information Processing Systems · 被引 368)
- NaVILA: Legged Robot Vision-Language-Action Model for Navigation (2024 · Robotics · 被引 243)
- NVILA: Efficient Frontier Visual Language Models (2024 · Computer Vision and Pattern Recognition · 被引 229)
- EgoVLA: Learning Vision-Language-Action Models from Egocentric Human Videos (2025 · arXiv.org · 被引 122)
- RGBD Objects in the Wild: Scaling Real-World 3D Object Learning from RGB-D Videos (2024 · Computer Vision and Pattern Recognition · 被引 112)
- RegionGPT: Towards Region Understanding Vision Language Model (2024 · Computer Vision and Pattern Recognition · 被引 102)