P
PaperPicks
Conferences
Andrew Zisserman
University of Oxford, UK
31 papers at tracked venues · 23 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0002-8945-8573 ↗
Google Scholar ↗
Homepage ↗
Venues
CVPR
×9
ICCV
×5
NeurIPS
×5
ECCV
×4
MICCAI
×3
AAAI
×1
ACL
×1
ACM MM
×1
InterSpeech
×1
SIGIR
×1
Frequent coauthors
Niki Amini-Naieni
DBLP profile ↗
ORCID search ↗
×2
Ragav Sachdeva
DBLP profile ↗
ORCID search ↗
×2
Tengda Han
DBLP profile ↗
ORCID search ↗
×2
Junyu Xie
DBLP profile ↗
ORCID search ↗
×2
Guanqi Zhan
DBLP profile ↗
ORCID search ↗
×2
Zifan Jiang
DBLP profile ↗
ORCID search ↗
×1
Prasanna Sridhar
DBLP profile ↗
ORCID search ↗
×1
Owen Pullen
DBLP profile ↗
ORCID search ↗
×1
Zhongrui Gui
DBLP profile ↗
ORCID search ↗
×1
Piyush Bagad
DBLP profile ↗
ORCID search ↗
×1
Goker Erdogan
DBLP profile ↗
ORCID search ↗
×1
Youngjoon Jang
DBLP profile ↗
ORCID search ↗
×1
Papers
Open-World Object Counting in Videos
AAAI 2026
·
Niki Amini-Naieni
DBLP profile ↗
ORCID search ↗
Segment, Embed, and Align: A Universal Recipe for Aligning Subtitles to Signing
ACL 2026
·
Zifan Jiang
DBLP profile ↗
ORCID search ↗
WISE: A Multimodal Search Engine for Visual Scenes, Audio, Objects, Faces, Speech, and Metadata
SIGIR 2026
·
Prasanna Sridhar
DBLP profile ↗
ORCID search ↗
A Simple Modality-Agnostic Representation for Scoliosis Phenotyping
MICCAI 2025
·
Owen Pullen
DBLP profile ↗
ORCID search ↗
Character-Centric Understanding of Animated Movies
ACM MM 2025
·
Zhongrui Gui
DBLP profile ↗
ORCID search ↗
Chirality in Action: Time-Aware Video Representation Learning by Latent Straightening
NeurIPS 2025
·
Piyush Bagad
DBLP profile ↗
ORCID search ↗
From Panels to Prose: Generating Literary Narratives from Comics
ICCV 2025
·
Ragav Sachdeva
DBLP profile ↗
ORCID search ↗
LayerLock: Non-Collapsing Representation Learning with Progressive Freezing
ICCV 2025
·
Goker Erdogan
DBLP profile ↗
ORCID search ↗
Learning from Streaming Video with Orthogonal Gradients
CVPR 2025
·
Tengda Han
DBLP profile ↗
ORCID search ↗
Lost in Translation, Found in Context: Sign Language Translation with Contextual Cues
CVPR 2025
·
Youngjoon Jang
DBLP profile ↗
ORCID search ↗
SciVid: Cross-Domain Evaluation of Video Models in Scientific Applications
ICCV 2025
·
Yana Hasson
DBLP profile ↗
ORCID search ↗
Shot-by-Shot: Film-Grammar-Aware Training-Free Audio Description Generation
ICCV 2025
·
Junyu Xie
DBLP profile ↗
ORCID search ↗
Understanding Co-Speech Gestures in-the-Wild
ICCV 2025
·
Sindhu B. Hegde
DBLP profile ↗
ORCID search ↗
3D Spine Shape Estimation from Single 2D DXA
MICCAI 2024
·
Emmanuelle Bourigault
DBLP profile ↗
ORCID search ↗
A General Protocol to Probe Large Vision Models for 3D Physical Understanding
NeurIPS 2024
·
Guanqi Zhan
DBLP profile ↗
ORCID search ↗
A Simple Recipe for Contrastively Pre-Training Video-First Encoders Beyond 16 Frames
CVPR 2024
·
Pinelopi Papalampidi
DBLP profile ↗
ORCID search ↗
Amodal Ground Truth and Completion in the Wild
CVPR 2024
·
Guanqi Zhan
DBLP profile ↗
ORCID search ↗
Appearance-Based Refinement for Object-Centric Motion Segmentation
ECCV 2024
·
Junyu Xie
DBLP profile ↗
ORCID search ↗
AutoAD III: The Prequel - Back to the Pixels
CVPR 2024
·
Tengda Han
DBLP profile ↗
ORCID search ↗
Automated Spinal MRI Labelling from Reports Using a Large Language Model
MICCAI 2024
·
Robin Y. Park
DBLP profile ↗
ORCID search ↗
CountGD: Multi-Modal Open-World Counting
NeurIPS 2024
·
Niki Amini-Naieni
DBLP profile ↗
ORCID search ↗
FlexCap: Describe Anything in Images in Controllable Detail
NeurIPS 2024
·
Debidatta Dwibedi
DBLP profile ↗
ORCID search ↗
Learning from One Continuous Video Stream
CVPR 2024
·
João Carreira
DBLP profile ↗
ORCID search ↗
Made to Order: Discovering Monotonic Temporal Changes via Self-supervised Video Ordering
ECCV 2024
·
Charig Yang
DBLP profile ↗
ORCID search ↗
N2F2: Hierarchical Scene Understanding with Nested Neural Feature Fields
ECCV 2024
·
Yash Bhalgat
DBLP profile ↗
ORCID search ↗
Separating the "Chirp" from the "Chat": Self-supervised Visual Grounding of Sound and Language
CVPR 2024
·
Mark Hamilton
DBLP profile ↗
ORCID search ↗
Speech Recognition Models are Strong Lip-readers
InterSpeech 2024
·
K. R. Prajwal
DBLP profile ↗
ORCID search ↗
TAPVid-3D: A Benchmark for Tracking Any Point in 3D
NeurIPS 2024
·
Skanda Koppula
DBLP profile ↗
ORCID search ↗
TIM: A Time Interval Machine for Audio-Visual Action Recognition
CVPR 2024
·
Jacob Chalk
DBLP profile ↗
ORCID search ↗
Text-Conditioned Resampler For Long Form Video Understanding
ECCV 2024
·
Bruno Korbar
DBLP profile ↗
ORCID search ↗
The Manga Whisperer: Automatically Generating Transcriptions for Comics
CVPR 2024
·
Ragav Sachdeva
DBLP profile ↗
ORCID search ↗