Zhuoyi Yang
发表论文 8 篇 · 总被引 2057 次 · h-index 6
代表论文
- CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer (2024 · International Conference on Learning Representations · 被引 2286)
- Concat-ID: Towards Universal Identity-Preserving Video Synthesis (2025 · 2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW) · 被引 31)
- VPO: Aligning Text-to-Video Generation Models with Prompt Optimization (2025 · IEEE International Conference on Computer Vision · 被引 24)
- Kaleido: Open-Sourced Multi-Subject Reference Video Generation Model (2025 · arXiv.org · 被引 15)
- SCAIL: Towards Studio-Grade Character Animation via In-Context Learning of 3D-Consistent Pose Representations (2025 · arXiv.org · 被引 11)
- Delving into Latent Spectral Biasing of Video VAEs for Superior Diffusability (2025 · arXiv.org · 被引 9)