PPaperPicks

Jianfeng Dong

23 papers at tracked venues · 19 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. FashionMAC: Deformation-Free Fashion Image Generation with Fine-Grained Model Appearance Customization
  2. Advancing Ship Re-Identification in the Wild: The ShipReID-2400 Benchmark Dataset and D2InterNet Baseline Method
  3. Audio Does Matter: Importance-Aware Multi-Granularity Fusion for Video Moment Retrieval
  4. Cooperative or Competitive? Understanding the Interaction between Attention Heads From A Game Theory Perspective
  5. Dynamic Adapter with Semantics Disentangling for Cross-lingual Cross-modal Retrieval
  6. Enhancing Partially Relevant Video Retrieval with Robust Alignment Learning
  7. Fit the Distribution: Cross-Image/Prompt Adversarial Attacks on Multimodal Large Language Models
  8. IVCR-200K: A Large-Scale Multi-turn Dialogue Benchmark for Interactive Video Corpus Retrieval
  9. LLM-Assisted Entropy-Based Adaptive Distillation for Unsupervised Fine-Grained Visual Representation Learning
    ICCV 2025 · Jianfeng Dong
  10. Multi-Pair Temporal Sentence Grounding via Multi-Thread Knowledge Transfer Network
  11. Multimodal Language Models See Better When They Look Shallower
  12. Open-World Fine-Grained Fashion Retrieval with LLM-based Commonsense Knowledge Infusion
    SIGIR 2025 · Jianfeng Dong
  13. Towards Building Model/Prompt-Transferable Attackers against Large Vision-Language Models
  14. Towards Efficient General Feature Prediction in Masked Skeleton Modeling
  15. Towards Ship License Plate Recognition in the Wild: A Large Benchmark and Strong Baseline
  16. CL2CM: Improving Cross-Lingual Cross-Modal Retrieval via Cross-Lingual Knowledge Transfer
  17. Frequency-Aware GAN for Imperceptible Transfer Attack on 3D Point Clouds
  18. Let All Be Whitened: Multi-Teacher Distillation for Efficient Visual Retrieval
  19. Not All Inputs Are Valid: Towards Open-Set Video Moment Retrieval using Language
  20. Rethinking Video Deblurring with Wavelet-Aware Dynamic Transformer and Diffusion Model
  21. Rethinking Weakly-Supervised Video Temporal Grounding From a Game Perspective
  22. Temporal Sentence Grounding with Relevance Feedback in Videos
    NeurIPS 2024 · Jianfeng Dong
  23. Unsupervised Domain Adaptative Temporal Sentence Localization with Mutual Information Maximization