PPaperPicks

Wanlong Fang

12 papers at tracked venues · 11 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Disentangling Adversarial Prompts: A Semantic-Graph Defense for Robust LLM Security
  2. Rethinking Video-Language Model from the Language Input Perspective
  3. To Align or Not to Align: Strategic Multimodal Representation Alignment for Optimal Performance
    AAAI 2026 · Wanlong Fang
  4. Towards Unified Vision-Language Models with Incomplete Multi-Modal Inputs
  5. Unveiling the Fragility of Vision-Language Models: Multi-Modal Adversarial Synergy via Texture-Constrained Perturbations and Cross-Modal Optimization
  6. Can LLMs Reason Over Non-Text Modalities in a Training-Free Manner? A Case Study with In-Context Representation Learning
  7. Hierarchical Semantic-Augmented Navigation: Optimal Transport and Graph-Driven Reasoning for Vision-Language Navigation
  8. Multi-Pair Temporal Sentence Grounding via Multi-Thread Knowledge Transfer Network
  9. Turing Patterns for Multimedia: Reaction-Diffusion Multi-Modal Fusion for Language-Guided Video Moment Retrieval
  10. Fewer Steps, Better Performance: Efficient Cross-Modal Clip Trimming for Video Moment Retrieval Using Language
  11. Not All Inputs Are Valid: Towards Open-Set Video Moment Retrieval using Language
  12. Rethinking Weakly-Supervised Video Temporal Grounding From a Game Perspective