Hanshen Zhu
发表论文 6 篇 · 总被引 28 次 · h-index 3
代表论文
- Benchmarking Vision-Language Models on Chinese Ancient Documents: From OCR to Knowledge Reasoning (2025 · arXiv.org · 被引 15)
- Training-Free Geometric Image Editing on Diffusion Models (2025 · IEEE International Conference on Computer Vision · 被引 11)
- TextPecker: Rewarding Structural Anomaly Quantification for Enhancing Visual Text Rendering (2026 · arXiv.org · 被引 9)
- SemiETS: Integrating Spatial and Content Consistencies for Semi-Supervised End-to-end Text Spotting (2025 · Computer Vision and Pattern Recognition · 被引 3)
- DenseAnnotate: Enabling Scalable Dense Caption Collection for Images and 3D Scenes via Spoken Descriptions (2025 · arXiv.org · 被引 1)
- SeFi-Image: A Text-to-Image Foundation Model with Semantic-First Diffusion (2026 · arXiv.org)