PPaperPicks

Alex C. Kot

Nanyang Technological University, Singapore

24 papers at tracked venues · 20 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. From Pretrain to Pain: Adversarial Vulnerability of Video Foundation Models Without Task Knowledge
  2. SAVER: Mitigating Hallucinations in Large Vision-Language Models via Style-Aware Visual Early Revision
  3. Backdoor Attacks Against No-Reference Image Quality Assessment Models via a Scalable Trigger
  4. MTL-UE: Learning to Learn Nothing for Multi-Task Learning
  5. Pay Attention to the Foreground in Object-Centric Learning
  6. Reconciling Stochastic and Deterministic Strategies for Zero-shot Image Restoration using Diffusion Model in Dual
  7. Temporal Unlearnable Examples: Preventing Personal Video Data from Unauthorized Exploitation by Object Tracking
  8. Theoretical Insights in Model Inversion Robustness and Conditional Entropy Maximization for Collaborative Inference Systems
  9. Vid-Group: Temporal Video Grounding Pretraining from Unlabeled Videos in the Wild
  10. BenchLMM: Benchmarking Cross-Style Visual Capability of Large Multimodal Models
  11. Compress Clean Signal from Noisy Raw Image: A Self-Supervised Approach
  12. ContextGS : Compact 3D Gaussian Splatting with Anchor Level Context Model
  13. Cross-Domain Few-Shot Segmentation via Iterative Support-Query Correspondence Mining
  14. E3M: Zero-Shot Spatio-Temporal Video Grounding with Expectation-Maximization Multimodal Modulation
  15. Evolving Storytelling: Benchmarks and Methods for New Character Customization with Diffusion Models
  16. From Chaos to Clarity: 3DGS in the Dark
  17. HideMIA: Hidden Wavelet Mining for Privacy-Enhancing Medical Image Analysis
  18. Local-Global Multi-Modal Distillation for Weakly-Supervised Temporal Video Grounding
  19. Omnipotent Distillation with LLMs for Weakly-Supervised Natural Language Video Localization: When Divergence Meets Consistency
  20. Purify Unlearnable Examples via Rate-Constrained Variational Autoencoders
  21. STSP: Spatial-Temporal Subspace Projection for Video Class-Incremental Learning
  22. SinSR: Diffusion-Based Image Super-Resolution in a Single Step
  23. Suppress and Rebalance: Towards Generalized Multi-Modal Face Anti-Spoofing
  24. Towards Physical World Backdoor Attacks Against Skeleton Action Recognition