Yuxin Xie
发表论文 25 篇 · 总被引 177 次 · h-index 8
代表论文
- VARGPT: Unified Understanding and Generation in a Visual Autoregressive Multimodal Large Language Model (2025 · arXiv.org · 被引 46)
- VARGPT-v1.1: Improve Visual Autoregressive Large Unified Model via Iterative Instruction Tuning and Reinforcement Learning (2025 · arXiv.org · 被引 23)
- SimTxtSeg: Weakly-Supervised Medical Image Segmentation with Simple Text Cues (2024 · International Conference on Medical Image Computing and Computer-Assisted Intervention · 被引 20)
- VASparse: Towards Efficient Visual Hallucination Mitigation via Visual-Aware Token Sparsification (2025 · Computer Vision and Pattern Recognition · 被引 19)
- HeartMuLa: A Family of Open Sourced Music Foundation Models (2026 · arXiv.org · 被引 13)
- Do we really have to filter out random noise in pre-training data for language models? (2025 · arXiv.org · 被引 13)