PPaperPicks

Nong Sang

25 papers at tracked venues · 24 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Learning to Tell Apart: Weakly Supervised Video Anomaly Detection via Disentangled Semantic Alignment
  2. Adaptive Prototype Replay for Class Incremental Semantic Segmentation
  3. CTR-Driven Advertising Image Generation with Multimodal Large Language Models
  4. Continual Gaussian Mixture Distribution Modeling for Class Incremental Semantic Segmentation
  5. DMPT: Decoupled Modality-Aware Prompt Tuning for Multi-Modal Object Re-Identification
  6. Holmes-VAU: Towards Long-term Video Anomaly Understanding at Any Granularity
  7. L-Man: A Large Multi-modal Model Unifying Human-centric Tasks
  8. MP-Mat: A 3D-and-Instance-Aware Human Matting and Editing Framework with Multiplane Representation
  9. NTIRE 2025 Challenge on Cross-Domain Few-Shot Object Detection: Methods and Results
  10. Partial Forward Blocking: A Novel Data Pruning Paradigm for Lossless Training Acceleration
  11. ReID5o: Achieving Omni Multi-modal Person Re-identification in a Single Model
  12. Structural Pruning via Spatial-aware Information Redundancy for Semantic Segmentation
  13. StyleSRN: Scene Text Image Super-Resolution with Text Style Embedding
  14. Towards Reliable and Holistic Visual In-Context Learning Prompt Selection
  15. VideoLucy: Deep Memory Backtracking for Long Video Understanding
  16. A Recipe for Scaling up Text-to-Video Generation with Text-free Videos
  17. Cross-video Identity Correlating for Person Re-identification Pre-training
  18. HR-Pro: Point-Supervised Temporal Action Localization via Hierarchical Reliability Propagation
  19. Hierarchical Spatio-temporal Decoupling for Text-to- Video Generation
  20. Open-Vocabulary Semantic Segmentation with Image Embedding Balancing
  21. PLIP: Language-Image Pre-training for Person Representation Learning
  22. Real-Time Exposure Correction via Collaborative Transformations and Adaptive Sampling
  23. SCTNet: Single-Branch CNN with Transformer Semantic Information for Real-Time Segmentation
  24. Tunnel Try-on: Excavating Spatial-temporal Tunnels for High-quality Virtual Try-on in Videos
  25. UFineBench: Towards Text-based Person Retrieval with Ultra-fine Granularity