PPaperPicks

Jungong Han

34 papers at tracked venues · 22 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Tracking and Segmenting Anything in Any Modality
  2. AdaTP: Attention-Debiased Token Pruning for Video Large Language Models
  3. Advancing Reliable Test-Time Adaptation of Vision-Language Models under Visual Variations
  4. DSMoE: Matrix-Partitioned Experts with Dynamic Routing for Computation-Efficient Dense LLMs
  5. DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval
  6. Extending LLM Context Window with Adaptive Grouped Positional Encoding: A Training-Free Method
  7. LSNet: See Large, Focus Small
  8. MCHM25: Multimedia Computing for Health and Medicine
  9. Mitigating Hallucinations in Multi-modal Large Language Models via Image Token Attention-Guided Decoding
  10. PrefixKV: Adaptive Prefix KV Cache is What Vision Instruction-Following Models Need for Efficient Generation
  11. Promptable Anomaly Segmentation with SAM Through Self-Perception Tuning
  12. Rethinking Score Distilling Sampling for 3D Editing and Generation
  13. Scaffold-BPE: Enhancing Byte Pair Encoding for Large Language Models with Simple and Effective Scaffold Token Removal
  14. Sequential Joint Dependency Aware Human Pose Estimation with State Space Model
  15. SimMLM: A Simple Framework for Multi-Modal Learning with Missing Modality
  16. Temporal Scaling Law for Large Language Models
  17. Unlocking the Potential of Diffusion Priors in Blind Face Restoration
  18. Unsupervised Patch-GAN with Targeted Patch Ranking for Fine-Grained Novelty Detection in Medical Imaging
  19. YOLOE: Real-Time Seeing Anythi
  20. Context Enhancement with Reconstruction as Sequence for Unified Unsupervised Anomaly Detection
  21. Eliminate Before Align: A Remote Sensing Image-Text Retrieval Framework with Keyword Explicit Reasoning
  22. FedGMKD: An Efficient Prototype Federated Learning Framework through Knowledge Distillation and Discrepancy-Aware Aggregation
  23. Learn from the Learnt: Source-Free Active Domain Adaptation via Contrastive Sampling and Visual Persistence
  24. On the Approximation Risk of Few-Shot Class-Incremental Learning
  25. One-dimensional Adapter to Rule Them All: Concepts, Diffusion Models and Erasing Applications
  26. PVUW 2024 Challenge on Complex Video Understanding: Methods and Results
  27. PYRA: Parallel Yielding Re-activation for Training-Inference Efficient Task Adaptation
  28. Pseudo-labelling Should Be Aware of Disguising Channel Activations
  29. Rep ViT: Revisiting Mobile CNN From ViT Perspective
  30. Revisiting motion information for RGB-Event tracking with MOT philosophy
  31. TaD: A Plug-and-Play Task-Aware Decoding Method to Better Adapt LLMs on Downstream Tasks
  32. The Second Visual Object Tracking Segmentation VOTS2024 Challenge Results
  33. WaveFace: Authentic Face Restoration with Efficient Frequency Recovery
  34. YOLOv10: Real-Time End-to-End Object Detection