Wenyi Hong
机构:Tsinghua University
发表论文 13 篇 · 总被引 270 次 · h-index 5
代表论文
- CogView2: Faster and Better Text-to-Image Generation via Hierarchical Transformers (2022 · arXiv (Cornell University) · 被引 121)
- CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers (2022 · arXiv (Cornell University) · 被引 116)
- MotionBench: Benchmarking and Improving Fine-Grained Video Motion Understanding for Vision Language Models (2025 · 被引 9)
- CogVLM2: Visual Language Models for Image and Video Understanding (2024 · arXiv (Cornell University) · 被引 7)
- CogAgent: A Visual Language Model for GUI Agents (2023 · arXiv (Cornell University) · 被引 6)
- Relay Diffusion: Unifying diffusion process across resolutions for image synthesis (2023 · arXiv (Cornell University) · 被引 5)