Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Hanshen Zhu

发表论文 6 篇 · 总被引 28 次 · h-index 3

代表论文

  • Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning (2025 · arXiv.org · 被引 15)
  • Training-Free Geometric Image Editing on Diffusion Models (2025 · IEEE International Conference on Computer Vision · 被引 11)
  • TextPecker: Rewarding Structural Anomaly Quantification for Enhancing Visual Text Rendering (2026 · arXiv.org · 被引 9)
  • SemiETS: Integrating Spatial and Content Consistencies for Semi-Supervised End-to-end Text Spotting (2025 · Computer Vision and Pattern Recognition · 被引 3)
  • DenseAnnotate: Enabling Scalable Dense Caption Collection for Images and 3D Scenes via Spoken Descriptions (2025 · arXiv.org · 被引 1)
  • SeFi-Image: A Text-to-Image Foundation Model with Semantic-First Diffusion (2026 · arXiv.org)