P
PaperPicks
Conferences
Yige Li
13 papers at tracked venues · 9 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
EMNLP
×3
ICLR
×2
ICML
×2
NeurIPS
×2
AAAI
×1
ACL
×1
CVPR
×1
EACL
×1
Frequent coauthors
Wei Zhao
DBLP profile ↗
ORCID search ↗
×3
Hanxun Huang
DBLP profile ↗
ORCID search ↗
×2
Yunhao Feng
DBLP profile ↗
ORCID search ↗
×1
Jiaming Zhang
DBLP profile ↗
ORCID search ↗
×1
Peihai Jiang
DBLP profile ↗
ORCID search ↗
×1
Yunhan Zhao
DBLP profile ↗
ORCID search ↗
×1
Nay Myat Min
DBLP profile ↗
ORCID search ↗
×1
Zhe Li
DBLP profile ↗
ORCID search ↗
×1
Shen Dong
DBLP profile ↗
ORCID search ↗
×1
Papers
BackdoorAgent: A Unified Framework for Backdoor Attacks on LLM-based Agents
ACL 2026
·
Yunhao Feng
DBLP profile ↗
ORCID search ↗
Unleashing the Unseen: Harnessing Benign Datasets for Jailbreaking Large Language Models
EACL 2026
·
Wei Zhao
DBLP profile ↗
ORCID search ↗
Anyattack: Towards Large-scale Self-supervised Adversarial Attacks on Vision-language Models
CVPR 2025
·
Jiaming Zhang
DBLP profile ↗
ORCID search ↗
Backdoor Token Unlearning: Exposing and Defending Backdoors in Pretrained Language Models
AAAI 2025
·
Peihai Jiang
DBLP profile ↗
ORCID search ↗
BackdoorLLM: A Comprehensive Benchmark for Backdoor Attacks and Defenses on Large Language Models
NeurIPS 2025
·
Yige Li
BlueSuffix: Reinforced Blue Teaming for Vision-Language Models Against Jailbreak Attacks
ICLR 2025
·
Yunhan Zhao
DBLP profile ↗
ORCID search ↗
CROW: Eliminating Backdoors from Large Language Models via Internal Consistency Regularization
ICML 2025
·
Nay Myat Min
DBLP profile ↗
ORCID search ↗
Detecting Backdoor Samples in Contrastive Language Image Pretraining
ICLR 2025
·
Hanxun Huang
DBLP profile ↗
ORCID search ↗
Do Influence Functions Work on Large Language Models?
EMNLP 2025
·
Zhe Li
DBLP profile ↗
ORCID search ↗
Memory Injection Attacks on LLM Agents via Query-Only Interaction
NeurIPS 2025
·
Shen Dong
DBLP profile ↗
ORCID search ↗
X-Transfer Attacks: Towards Super Transferable Adversarial Attacks on CLIP
ICML 2025
·
Hanxun Huang
DBLP profile ↗
ORCID search ↗
Zero-Shot Defense Against Toxic Images via Inherent Multimodal Alignment in LVLMs
EMNLP 2025
·
Wei Zhao
DBLP profile ↗
ORCID search ↗
Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing
EMNLP 2024
·
Wei Zhao
DBLP profile ↗
ORCID search ↗