Mustafa Shukor
发表论文 26 篇 · 总被引 1168 次 · h-index 16
代表论文
- SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics (2025 · arXiv.org · 被引 472)
- Rewarded soups: towards Pareto-optimal alignment by interpolating weights fine-tuned on diverse rewards (2023 · Neural Information Processing Systems · 被引 274)
- Multimodal Autoregressive Pre-training of Large Vision Encoders (2024 · Computer Vision and Pattern Recognition · 被引 119)
- What Makes Multimodal In-Context Learning Work? (2024 · 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) · 被引 66)
- Unified Model for Image, Video, Audio and Language Tasks (2023 · Trans. Mach. Learn. Res. · 被引 59)
- Scaling Laws for Optimal Data Mixtures (2025 · Neural Information Processing Systems · 被引 49)