Yuliang Liu
发表论文 50 篇 · 总被引 2026 次 · h-index 17
代表论文
- TextMonkey: An OCR-Free Large Multimodal Model for Understanding Document (2024 · IEEE Transactions on Pattern Analysis and Machine Intelligence · 被引 211)
- MTVQA: Benchmarking Multilingual Text-Centric Visual Question Answering (2024 · Annual Meeting of the Association for Computational Linguistics · 被引 87)
- MonkeyOCR: Document Parsing with a Structure-Recognition-Relation Triplet Paradigm (2025 · arXiv.org · 被引 77)
- Liquid: Language Models are Scalable Multi-modal Generators (2024 · arXiv.org · 被引 45)
- WildDoc: How Far Are We from Achieving Comprehensive and Robust Document Understanding in the Wild? (2025 · Conference on Empirical Methods in Natural Language Processing · 被引 27)
- OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models (2025 · IEEE Transactions on Pattern Analysis and Machine Intelligence · 被引 23)