PPaperPicks

Vishal M. Patel

Johns Hopkins University, Baltimore, MD, USA

59 papers at tracked venues · 34 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Addressing Data Scarcity in Materials Science Research with Deep Generative Models
  2. DiffRegCD: Integrated Registration and Change Detection with Diffusion Features
  3. F-ViTA: Foundation Model Guided Visible-to-Infrared Translation
  4. GHOST: Getting to the Bottom of Hallucinations with A Multi-round Consistency Benchmark
  5. Morphing Through Time: Diffusion-Based Bridging of Temporal Gaps for Robust Alignment in Change Detection
  6. ProCrop: Learning Aesthetic Image Cropping from Professional Compositions
  7. Referring Change Detection in Remote Sensing Imagery
  8. A Mamba-Based Siamese Network for Remote Sensing Change Detection
  9. A Technical Report on "Erasing the Invisible": The 2024 NeurIPS Competition on Stress Testing Image Watermarks
  10. AWRaCLe: All-Weather Image Restoration Using Visual In-Context Learning
  11. Active Learning for Vision-Language Models
  12. DDPM-CD: Denoising Diffusion Probabilistic Models as Feature Extractors for Remote Sensing Change Detection
    WACV 2025 ·
    Wele Gedara Chaminda Bandara
  13. Deep Metric Learning for Unsupervised Remote Sensing Change Detection
    WACV 2025 ·
    Wele Gedara Chaminda Bandara
  14. Distilling Multi-modal Large Language Models for Autonomous Driving
  15. FaceXFormer: A Unified Transformer for Facial Analysis
  16. Field-DiT: Diffusion Transformer on Unified Video, 3D, and Game Field Generation
  17. Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning
  18. Frame by Familiar Frame: Understanding Replication in Video Diffusion Models
  19. GenDeg: Diffusion-based Degradation Synthesis for Generalizable All-In-One Image Restoration
  20. Harmonyseg: Tubular Structure Segmentation With Deep-Shallow Feature Fusion and Growth-Suppression Balanced Loss
  21. Improving Conditional Diffusion Models through Re-Noising from Unconditional Diffusion Priors
  22. Low-Rank Adaptation-Based All-Weather Removal for Autonomous Navigation
  23. Lux Post Facto: Learning Portrait Performance Relighting with Conditional Video Diffusion and a Hybrid Dataset
  24. MIRE: Matched Implicit Neural Representations
  25. MambaRecon: MRI Reconstruction with Structured State Space Models
  26. Multimodal 3D Object Detection on Unseen Domains
  27. Open Vision Reasoner: Transferring Linguistic Cognitive Behavior for Visual Reasoning
  28. PETALface: Parameter Efficient Transfer Learning for Low-Resolution Face Recognition
  29. PIN: Prolate Spheroidal Wave Function-based Implicit Neural Representations
  30. Perception in Reflection
  31. ReBotNet: Fast Real-Time Video Enhancement
    WACV 2025 ·
    Jeya Maria Jose Valanarasu
  32. SINR: Sparsity Driven Compressed Implicit Neural Representations
  33. STEREO: A Two-Stage Framework for Adversarially Robust Concept Erasing from Text-to-Image Diffusion Models
  34. Scaling Transformer-Based Novel View Synthesis with Models Token Disentanglement and Synthetic Data
    ICCV 2025 ·
    Nithin Gopalakrishnan Nair
  35. SegFace: Face Segmentation of Long-Tail Classes
  36. StepAL: Step-Aware Active Learning for Cataract Surgical Videos
  37. SyncNoise: Geometrically Consistent Noise Prediction for Instruction-based 3D Editing
  38. The Power of Context: How Multimodality Improves Image Super-Resolution
  39. Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
  40. UniRes: Universal Image Restoration for Complex Degradations
  41. Black-Box Adaptation for Medical Image Segmentation
  42. CoDi: Conditional Diffusion Distillation for Higher-Fidelity and Faster Image Generation
  43. CrowdDiff: Multi-Hypothesis Crowd Density Estimation Using Diffusion Models
  44. Dense Multimodal Alignment for Open-Vocabulary 3D Scene Understanding
  45. Entropic Open-Set Active Learning
  46. Equivariant Spatio-temporal Self-supervision for LiDAR Object Detection
  47. Federated Black-Box Adaptation for Semantic Segmentation
  48. Gradient-Regularized Out-of-Distribution Detection
  49. Holo-Relighting: Controllable Volumetric Portrait Relighting from a Single Image
  50. JeDi: Joint-Image Diffusion Models for Finetuning-Free Personalized Text-to-Image Generation
  51. LQMFormer: Language-Aware Query Mask Transformer for Referring Image Segmentation
  52. Leveraging Thermal Modality to Enhance Reconstruction in Low-Light Conditions
  53. MaxFusion: Plug&Play Multi-modal Generation in Text-to-Image Diffusion Models
    ECCV 2024 ·
    Nithin Gopalakrishnan Nair
  54. ModelMix: A New Model-Mixup Strategy to Minimize Vicinal Risk Across Tasks for Few-Scribble Based Cardiac Segmentation
  55. MonoDiff: Monocular 3D Object Detection and Pose Estimation with Diffusion Models
  56. ReGS: Reference-based Controllable Scene Stylization with Gaussian Splatting
  57. S-SAM: SVD-Based Fine-Tuning of Segment Anything Model for Medical Image Segmentation
  58. View-decoupled Transformer for Person Re-identification under Aerial-ground Camera Network
  59. Wild-GS: Real-Time Novel View Synthesis from Unconstrained Photo Collections