Chenming Zhu
发表论文 17 篇 · 总被引 829 次 · h-index 13
代表论文
- LLaVA-3D: A Simple Yet Effective Pathway to Empowering LMMs with 3D Capabilities (2024 · IEEE International Conference on Computer Vision · 被引 197)
- EmbodiedScan: A Holistic Multi-Modal 3D Perception Suite Towards Embodied AI (2023 · Computer Vision and Pattern Recognition · 被引 189)
- MMSI-Bench: A Benchmark for Multi-Image Spatial Intelligence (2025 · arXiv.org · 被引 172)
- StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling (2025 · arXiv.org · 被引 118)
- ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities (2024 · European Conference on Computer Vision · 被引 64)
- LLaVA-3D: A Simple yet Effective Pathway to Empowering LMMs with 3D-awareness (2024 · arXiv.org · 被引 60)