Ronghang Hu
发表论文 48 篇 · 总被引 7031 次 · h-index 26
代表论文
- ConvNeXt V2: Co-designing and Scaling ConvNets with Masked Autoencoders (2023 · 被引 1448)
- FLAVA: A Foundational Language And Vision Alignment Model (2022 · 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) · 被引 511)
- UniT: Multimodal Multitask Learning with a Unified Transformer (2021 · 2021 IEEE/CVF International Conference on Computer Vision (ICCV) · 被引 283)
- SAM 2: Segment Anything in Images and Videos (2024 · arXiv (Cornell University) · 被引 260)
- Scaling Language-Image Pre-Training via Masking (2023 · 被引 225)
- Iterative Answer Prediction With Pointer-Augmented Multimodal Transformers for TextVQA (2020 · 被引 217)