P
PaperPicks
Conferences
Zhenhong Zhou
18 papers at tracked venues · 13 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
ACL
×8
EMNLP
×5
AAAI
×2
ICLR
×1
ICML
×1
NeurIPS
×1
Frequent coauthors
Yuanhe Zhang
DBLP profile ↗
ORCID search ↗
×3
Liang Lin
DBLP profile ↗
ORCID search ↗
×2
Jin Wang
DBLP profile ↗
ORCID search ↗
×1
Yibo Zhang
DBLP profile ↗
ORCID search ↗
×1
Yu Jiang
DBLP profile ↗
ORCID search ↗
×1
Zixuan Wang
DBLP profile ↗
ORCID search ↗
×1
Pengyu Zhu
DBLP profile ↗
ORCID search ↗
×1
Wei Zhang
DBLP profile ↗
ORCID search ↗
×1
Zherui Li
DBLP profile ↗
ORCID search ↗
×1
Quan Liu
DBLP profile ↗
ORCID search ↗
×1
Rongwu Xu
DBLP profile ↗
ORCID search ↗
×1
Papers
Backdoor Collapse: Eliminating Unknown Threats Via Known Backdoor Aggregation In Language Models
ACL 2026
·
Liang Lin
DBLP profile ↗
ORCID search ↗
CORBA: Contagious Recursive Blocking Attacks on Multi-Agent Systems Based on Large Language Models
ACL 2026
·
Zhenhong Zhou
HearSay Benchmark: Do Audio LLMs Leak What They Hear?
ACL 2026
·
Jin Wang
DBLP profile ↗
ORCID search ↗
Hidden in the Noise: Unveiling Backdoors in Audio LLMs Alignment Through Latent Acoustic Pattern Triggers
AAAI 2026
·
Liang Lin
DBLP profile ↗
ORCID search ↗
RSA-Bench: Benchmarking Audio Large Models in Real-World Acoustic Scenarios
ACL 2026
·
Yibo Zhang
DBLP profile ↗
ORCID search ↗
RiskLab: A Controlled Toolkit for Probing Emergent Risks in LLM-Based Multi-Agent Systems
ACL 2026
·
Yu Jiang
DBLP profile ↗
ORCID search ↗
SEE: Signal Embedding Energy for Quantifying Noise Interference in Large Audio Language Models
ACL 2026
·
Yuanhe Zhang
DBLP profile ↗
ORCID search ↗
X-Router: Decoupling Knowledge and Reasoning for Cost-Effective LLM Inference
ACL 2026
·
Zixuan Wang
DBLP profile ↗
ORCID search ↗
Crabs: Consuming Resource via Auto-generation for LLM-DoS Attack under Black-box Settings
ACL 2025
·
Yuanhe Zhang
DBLP profile ↗
ORCID search ↗
DemonAgent: Dynamically Encrypted Multi-Backdoor Implantation Attack on LLM-based Agent
EMNLP 2025
·
Pengyu Zhu
DBLP profile ↗
ORCID search ↗
LIFEBENCH: Evaluating Length Instruction Following in Large Language Models
NeurIPS 2025
·
Wei Zhang
DBLP profile ↗
ORCID search ↗
On the Role of Attention Heads in Large Language Model Safety
ICLR 2025
·
Zhenhong Zhou
PD³F: A Pluggable and Dynamic DoS-Defense Framework against resource consumption attacks targeting Large Language Models
EMNLP 2025
·
Yuanhe Zhang
DBLP profile ↗
ORCID search ↗
Reinforced Lifelong Editing for Language Models
ICML 2025
·
Zherui Li
DBLP profile ↗
ORCID search ↗
Alignment-Enhanced Decoding: Defending Jailbreaks via Token-Level Adaptive Refining of Probability Distributions
EMNLP 2024
·
Quan Liu
DBLP profile ↗
ORCID search ↗
Course-Correction: Safety Alignment Using Synthetic Preferences
EMNLP 2024
·
Rongwu Xu
DBLP profile ↗
ORCID search ↗
How Alignment and Jailbreak Work: Explain LLM Safety through Intermediate Hidden States
EMNLP 2024
·
Zhenhong Zhou
Quantifying and Analyzing Entity-Level Memorization in Large Language Models
AAAI 2024
·
Zhenhong Zhou