Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Yuliang Liu

发表论文 50 篇 · 总被引 2026 次 · h-index 17

代表论文

  • TextMonkey: An OCR-Free Large Multimodal Model for Understanding Document (2024 · IEEE Transactions on Pattern Analysis and Machine Intelligence · 被引 211)
  • MTVQA: Benchmarking Multilingual Text-Centric Visual Question Answering (2024 · Annual Meeting of the Association for Computational Linguistics · 被引 87)
  • MonkeyOCR: Document Parsing with a Structure-Recognition-Relation Triplet Paradigm (2025 · arXiv.org · 被引 77)
  • Liquid: Language Models are Scalable Multi-modal Generators (2024 · arXiv.org · 被引 45)
  • WildDoc: How Far Are We from Achieving Comprehensive and Robust Document Understanding in the Wild? (2025 · Conference on Empirical Methods in Natural Language Processing · 被引 27)
  • OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models (2025 · IEEE Transactions on Pattern Analysis and Machine Intelligence · 被引 23)