PPaperPicks

Luc Van Gool

INSAIT, Sofia University "St. Kliment Ohridski", Sofia, Bulgaria

86 papers at tracked venues · 55 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Any Detector Can Detect Anything
  2. Autonomous Vehicle Path Planning by Searching with Differentiable Simulation
  3. EgoCross: Benchmarking Multimodal Large Language Models for Cross-Domain Egocentric Video Question Answering
  4. Optimizing against Infeasible Inclusions from Data for Semantic Segmentation through Morphology
  5. Unlocking Efficient Vehicle Dynamics Modeling via Analytic World Models
  6. Articulate3D: Holistic Understanding of 3D Scenes as Universal Scene Description
  7. Autonomous Vehicle Controllers From End-to-End Differentiable Simulation
  8. Benchmarking Multi-modal Semantic Segmentation under Sensor Failures: Missing and Noisy Modality Robustness
  9. Camera-Only 3D Panoptic Scene Completion for Autonomous Driving through Differentiable Object Shapes
  10. Diffusion-Based Particle-DETR for BEV Perception
  11. Domain-RAG: Retrieval-Guided Compositional Image Generation for Cross-Domain Few-Shot Object Detection
  12. Exploration-Driven Generative Interactive Environments
  13. Fine-Grained Spatial and Verbal Losses for 3D Visual Grounding
  14. Incremental Object Detection with Prompt-Based Methods
  15. LangHOPS: Language Grounded Hierarchical Open-Vocabulary Part Segmentation
  16. Language-Guided Instance-Aware Domain-Adaptive Panoptic Segmentation
  17. Learning to Prompt with Text Only Supervision for Vision-Language Models
  18. Leveraging Driver Field-of-View for Multimodal Ego-Trajectory Prediction
  19. Locate Anything on Earth: Advancing Open-Vocabulary Object Detection for Remote Sensing Community
  20. Low-Light Image Enhancement Using Event-Based Illumination Estimation
  21. MARS2 2025 Challenge on Multimodal Reasoning: Datasets, Methods, Results, Discussion, and Outlook
  22. MOBIUS: Big-to-Mobile Universal Instance Segmentation via Multi-modal Bottleneck Fusion and Calibrated Decoder Pruning
  23. MarkushGrapher: Joint Visual and Textual Recognition of Markush Structures
  24. Model Weights Reflect a Continuous Space of Input Image Domains
  25. NTIRE 2025 Challenge on Cross-Domain Few-Shot Object Detection: Methods and Results
  26. NTIRE 2025 Challenge on Event-Based Image Deblurring: Methods and Results
  27. ObjectRelator: Enabling Cross-View Object Relation Understanding Across Ego-Centric and Exo-Centric Perspectives
  28. One2Any: One-Reference 6D Pose Estimation for Any Object
  29. PBR-NeRF: Inverse Rendering with Physics-Based Neural Fields
  30. ReVLA: Reverting Visual Domain Limitation of Robotic Foundation Models
  31. Reducing Unimodal Bias in Multi-Modal Semantic Segmentation With Multi-Scale Functional Entropy Regularization
  32. Samba: Synchronized Set-of-Sequences Modeling for Multiple Object Tracking
  33. SceneSplat++: A Large Dataset and Comprehensive Benchmark for Language Gaussian Splatting
  34. SceneSplat: Gaussian Splatting-Based Scene Understanding with Vision-Language Pretraining
  35. Splat-SLAM: Globally Optimized RGB-only SLAM with 3D Gaussians
  36. StateSpaceDiffuser: Bringing Long Context to Diffusion World Models
  37. Sun Off, Lights on: Photorealistic Monocular Nighttime Simulation for Robust Semantic Perception
    WACV 2025 ·
    Konstantinos Tzevelekakis
  38. The Tenth NTIRE 2025 Image Denoising Challenge Report
  39. Understanding Museum Exhibits using Vision-Language Reasoning
  40. UniK3D: Universal Camera Monocular 3D Estimation
  41. What You Have is What You Track: Adaptive and Robust Multimodal Tracking
  42. XTrack: Multimodal Training Boosts RGB-X Video Object Trackers
  43. A Unified and Interpretable Emotion Representation and Expression Generation
  44. Bayesian Self-training for Semi-supervised 3D Segmentation
  45. Continuous Pose for Monocular Cameras in Neural Implicit Representation
  46. Cross-Domain Few-Shot Object Detection via Enhanced Open-Set Object Detector
  47. DGInStyle: Domain-Generalizable Semantic Segmentation with Image Diffusion Models and Stylized Semantic Control
  48. Deep Equilibrium Diffusion Restoration with Parallel Sampling
  49. Equivariant Multi-Modality Image Fusion
  50. Event-Free Moving Object Segmentation from Moving Ego Vehicle
  51. Four Ways to Improve Verbo-visual Fusion for Dense 3D Visual Grounding
  52. HandDiff: 3D Hand Pose Estimation with Diffusion on Image-Point Cloud
  53. I-Design: Personalized LLM Interior Designer
  54. Image Fusion via Vision-Language Model
  55. Implicit Zoo: A Large-Scale Dataset of Neural Implicit Functions for 2D Images and 3D Scenes
  56. Investigating the Effectiveness of Cross-Attention to Unlock Zero-Shot Editing of Text-to-Video Diffusion Models
  57. Know Your Neighbors: Improving Single-View Reconstruction via Spatial Vision-Language Reasoning
  58. Lego: Learning to Disentangle and Invert Personalized Concepts Beyond Object Appearance in Text-to-Image Diffusion Models
  59. Lightweight Image Super-Resolution via Flexible Meta Pruning
  60. Loopy-SLAM: Dense Neural SLAM with Loop Closures
  61. MICDrop: Masking Image and Depth Features via Complementary Dropout for Domain-Adaptive Semantic Segmentation
  62. MUSES: The Multi-sensor Semantic Perception Dataset for Driving Under Uncertainty
  63. Matching Anything by Segmenting Anything
  64. MoVideo: Motion-Aware Video Generation with Diffusion Model
  65. PALM: Predicting Actions through Language Models
  66. Probabilistic Sampling of Balanced K-Means using Adiabatic Quantum Computing
  67. Prompting Diffusion Representations for Cross-Domain Semantic Segmentation
  68. ROMEO: Revisiting Optimization Methods for Reconstructing 3D Human-Object Interaction Models From Images
  69. Real-World Mobile Image Denoising Dataset with Efficient Baselines
  70. Rethinking Few-shot 3D Point Cloud Semantic Segmentation
  71. SILC: Improving Vision Language Pretraining with Self-distillation
  72. SLAck: Semantic, Location, and Appearance Aware Open-Vocabulary Tracking
  73. Self-supervised Shape Completion via Involution and Implicit Correspondences
  74. SemiVL: Semi-Supervised Semantic Segmentation with Vision-Language Guidance
  75. Sharing Key Semantics in Transformer Makes Efficient Image Restoration
  76. Single-Model and Any-Modality for Video Object Tracking
  77. Stereo Risk: A Continuous Modeling Approach to Stereo Matching
  78. Summarize the Past to Predict the Future: Natural Language Descriptions of Context Boost Multimodal Object Interaction Anticipation
  79. Taming CLIP for Fine-Grained and Structured Visual Understanding of Museum Exhibits
  80. Ternary-Type Opacity and Hybrid Odometry for RGB NeRF-SLAM
  81. The BRAVO Semantic Segmentation Challenge Results in UNCV2024
  82. Towards Online Real-Time Memory-based Video Inpainting Transformers
  83. UniDepth: Universal Monocular Metric Depth Estimation
  84. Vanishing-Point-Guided Video Semantic Segmentation of Driving Scenes
  85. Video Foundation Model for Medical 3D Segmentation
  86. Walker: Self-supervised Multiple Object Tracking by Walking on Temporal Appearance Graphs