PPaperPicks

Lianli Gao

20 papers at tracked venues · 17 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Debiased Orthogonal Boundary-Driven Efficient Noise Mitigation
  2. AICL: Action In-Context Learning for Text-to-Video Generation
  3. DFDNet: Disentangling and Filtering Dynamics for Enhanced Video Prediction
  4. FlexAC: Towards Flexible Control of Associative Reasoning in Multimodal Large Language Models
  5. MMEvol: Empowering Multimodal Large Language Models with Evol-Instruct
  6. OmniCharacter: Towards Immersive Role-Playing Agents with Seamless Speech-Language Personality Interaction
  7. Safe + Safe = Unsafe? Exploring How Safe Images Can Be Exploited to Jailbreak Large Vision-Language Models
  8. SafePTR: Token-Level Jailbreak Defense in Multimodal LLMs via Prune-then-Restore Mechanism
  9. Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters Themselves
  10. Unlocking Smarter Device Control: Foresighted Planning with a World Model-Driven Code Execution Approach
  11. Alleviating Hallucinations in Large Vision-Language Models through Hallucination-Induced Optimization
  12. Any Target Can be Offense: Adversarial Example Generation via Generalized Latent Infection
  13. CoIN: A Benchmark of Continual Instruction Tuning for Multimodel Large Language Models
  14. DePT: Decoupled Prompt Tuning
  15. F³-Pruning: A Training-Free and Generalized Pruning Strategy towards Faster and Finer Text-to-Video Synthesis
  16. MPT: Multi-grained Prompt Tuning for Text-Video Retrieval
  17. MagicVFX: Visual Effects Synthesis in Just Minutes
  18. ProS: Prompting-to-Simulate Generalized Knowledge for Universal Cross-Domain Retrieval
  19. RoScenes: A Large-Scale Multi-view 3D Dataset for Roadside Perception
  20. SI-BiViT: Binarizing Vision Transformers with Spatial Interaction