Scholay

学术搜索 · AI 审稿 · LaTeX 协作

I. Laptev

发表论文 193 篇 · 总被引 40218 次 · h-index 76

代表论文

  • Vid2Seq: Large-Scale Pretraining of a Visual Language Model for Dense Video Captioning (2023 · Computer Vision and Pattern Recognition · 被引 414)
  • Zero-Shot Video Question Answering via Frozen Bidirectional Language Models (2022 · Neural Information Processing Systems · 被引 301)
  • Think Global, Act Local: Dual-scale Graph Transformer for Vision-and-Language Navigation (2022 · Computer Vision and Pattern Recognition · 被引 295)
  • Language Conditioned Spatial Relation Reasoning for 3D Object Grounding (2022 · Neural Information Processing Systems · 被引 165)
  • Instruction-driven history-aware policies for robotic manipulations (2022 · Conference on Robot Learning · 被引 157)
  • TubeDETR: Spatio-Temporal Video Grounding with Transformers (2022 · Computer Vision and Pattern Recognition · 被引 141)