Renrui Zhang
机构:Chinese University of Hong Kong · ORCID:0000-0003-4503-5277
发表论文 65 篇 · 总被引 2843 次 · h-index 20
代表论文
- OneTracker: Unifying Visual Object Tracking with Foundation Models and Efficient Tuning (2024 · 被引 130)
- RenderOcc: Vision-Centric 3D Occupancy Prediction with 2D Rendering Supervision (2024 · 被引 67)
- Referred by Multi-Modality: A Unified Temporal Transformer for Video Object Segmentation (2024 · Proceedings of the AAAI Conference on Artificial Intelligence · 被引 36)
- LiDAR-LLM: Exploring the Potential of Large Language Models for 3D LiDAR Understanding (2025 · Proceedings of the AAAI Conference on Artificial Intelligence · 被引 33)
- LLaVA-OneVision: Easy Visual Task Transfer (2024 · arXiv (Cornell University) · 被引 32)
- LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models (2024 · arXiv (Cornell University) · 被引 27)