Shuangrui Ding
ORCID:0000-0001-7033-774X
发表论文 43 篇 · 总被引 318 次 · h-index 10
代表论文
- Image Compression for Machine and Human Vision with Spatial-Frequency Adaptation (2024 · Lecture notes in computer science · 被引 23)
- SAM 3: Segment Anything with Concepts (2025 · arXiv (Cornell University) · 被引 12)
- Dispider: Enabling Video LLMs with Active Real-Time Interaction via Disentangled Perception, Decision, and Reaction (2025 · 被引 9)
- SongComposer: A Large Language Model for Lyric and Melody Generation in Song Composition (2025 · 被引 6)
- OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding? (2025 · 被引 5)
- Betrayed by Attention: A Simple yet Effective Approach for Self-supervised Video Object Segmentation (2024 · Lecture notes in computer science · 被引 4)