Andrew Zisserman
发表论文 839 篇 · 总被引 331644 次 · h-index 191
代表论文
- WhisperX: Time-Accurate Speech Transcription of Long-Form Audio (2023 · Interspeech · 被引 606)
- Perception Test: A Diagnostic Benchmark for Multimodal Video Models (2023 · Neural Information Processing Systems · 被引 366)
- TAPIR: Tracking Any Point with per-frame Initialization and temporal Refinement (2023 · IEEE International Conference on Computer Vision · 被引 344)
- Temporal Alignment Networks for Long-term Video (2022 · Computer Vision and Pattern Recognition · 被引 121)
- Verbs in Action: Improving verb understanding in video-language models (2023 · IEEE International Conference on Computer Vision · 被引 102)
- Multi-Modal Classifiers for Open-Vocabulary Object Detection (2023 · International Conference on Machine Learning · 被引 74)