PPaperPicks

Zhenhong Zhou

18 papers at tracked venues · 13 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Backdoor Collapse: Eliminating Unknown Threats Via Known Backdoor Aggregation In Language Models
  2. CORBA: Contagious Recursive Blocking Attacks on Multi-Agent Systems Based on Large Language Models
    ACL 2026 · Zhenhong Zhou
  3. HearSay Benchmark: Do Audio LLMs Leak What They Hear?
  4. Hidden in the Noise: Unveiling Backdoors in Audio LLMs Alignment Through Latent Acoustic Pattern Triggers
  5. RSA-Bench: Benchmarking Audio Large Models in Real-World Acoustic Scenarios
  6. RiskLab: A Controlled Toolkit for Probing Emergent Risks in LLM-Based Multi-Agent Systems
  7. SEE: Signal Embedding Energy for Quantifying Noise Interference in Large Audio Language Models
  8. X-Router: Decoupling Knowledge and Reasoning for Cost-Effective LLM Inference
  9. Crabs: Consuming Resource via Auto-generation for LLM-DoS Attack under Black-box Settings
  10. DemonAgent: Dynamically Encrypted Multi-Backdoor Implantation Attack on LLM-based Agent
  11. LIFEBENCH: Evaluating Length Instruction Following in Large Language Models
  12. On the Role of Attention Heads in Large Language Model Safety
    ICLR 2025 · Zhenhong Zhou
  13. PD³F: A Pluggable and Dynamic DoS-Defense Framework against resource consumption attacks targeting Large Language Models
  14. Reinforced Lifelong Editing for Language Models
  15. Alignment-Enhanced Decoding: Defending Jailbreaks via Token-Level Adaptive Refining of Probability Distributions
  16. Course-Correction: Safety Alignment Using Synthetic Preferences
  17. How Alignment and Jailbreak Work: Explain LLM Safety through Intermediate Hidden States
    EMNLP 2024 · Zhenhong Zhou
  18. Quantifying and Analyzing Entity-Level Memorization in Large Language Models
    AAAI 2024 · Zhenhong Zhou