PPaperPicks

Daizong Liu

29 papers at tracked venues · 24 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. LearnerCoMPASS: Intelligent Tutoring System with Dynamic Cognitive Diagnosis and Multi-Model Path Planning
  2. Rethinking Video-Language Model from the Language Input Perspective
  3. Spatial-Spectral Homogeneous Attacks on Physical-World Large Vision-Language Models
    AAAI 2026 · Daizong Liu
  4. Towards Unified Vision-Language Models with Incomplete Multi-Modal Inputs
  5. Audio Does Matter: Importance-Aware Multi-Granularity Fusion for Video Moment Retrieval
  6. Cooperative or Competitive? Understanding the Interaction between Attention Heads From A Game Theory Perspective
  7. Fast3D: Accelerating 3D Multi-modal Large Language Models for Efficient 3D Scene Understanding
  8. Fit the Distribution: Cross-Image/Prompt Adversarial Attacks on Multimodal Large Language Models
  9. Imperceptible 3D Point Cloud Attacks on Lattice-based Barycentric Coordinates
  10. LLM-Assisted Entropy-Based Adaptive Distillation for Unsupervised Fine-Grained Visual Representation Learning
  11. Learning from Few Samples: A Novel Approach for High-Quality Malcode Generation
  12. Misalignment Attack on Text-to-Image Models via Text Embedding Optimization and Inversion
  13. Multi-Pair Temporal Sentence Grounding via Multi-Thread Knowledge Transfer Network
  14. Open-World Fine-Grained Fashion Retrieval with LLM-based Commonsense Knowledge Infusion
  15. Seeing is Not Believing: Adversarial Natural Object Optimization for Hard-Label 3D Scene Attacks
    CVPR 2025 · Daizong Liu
  16. Towards Building Model/Prompt-Transferable Attackers against Large Vision-Language Models
  17. Advancing 3D Object Grounding Beyond a Single 3D Scene
  18. Cross-Task Knowledge Transfer for Semi-supervised Joint 3D Grounding and Captioning
  19. Explicitly Perceiving and Preserving the Local Geometric Structures for 3D Point Cloud Attack
    AAAI 2024 · Daizong Liu
  20. FLAT: Flux-Aware Imperceptible Adversarial Attacks on 3D Point Clouds
  21. Fewer Steps, Better Performance: Efficient Cross-Modal Clip Trimming for Video Moment Retrieval Using Language
  22. Frequency-Aware GAN for Imperceptible Transfer Attack on 3D Point Clouds
  23. Hiding Imperceptible Noise in Curvature-Aware Patches for 3D Point Cloud Attack
  24. Manifold Constraints for Imperceptible Adversarial Attacks on Point Clouds
  25. Not All Inputs Are Valid: Towards Open-Set Video Moment Retrieval using Language
  26. Pandora's Box: Towards Building Universal Attackers against Real-World Large Vision-Language Models
    NeurIPS 2024 · Daizong Liu
  27. Rethinking Weakly-Supervised Video Temporal Grounding From a Game Perspective
  28. Temporal Sentence Grounding with Relevance Feedback in Videos
  29. Unsupervised Domain Adaptative Temporal Sentence Localization with Mutual Information Maximization
    AAAI 2024 · Daizong Liu