Shiyu Huang
发表论文 12 篇 · 总被引 6468 次 · h-index 8
代表论文
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities (2025 · arXiv.org · 被引 3882)
- CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer (2024 · International Conference on Learning Representations · 被引 2286)
- LVBench: An Extreme Long Video Understanding Benchmark (2024 · IEEE International Conference on Computer Vision · 被引 420)
- GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning (2025 · 被引 353)
- CogVLM2: Visual Language Models for Image and Video Understanding (2024 · arXiv.org · 被引 260)
- GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning (2025 · arXiv.org · 被引 209)