P
PaperPicks
Conferences
Zhexin Zhang
13 papers at tracked venues · 12 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
ACL
×7
ACM MM
×2
AAAI
×1
CHI
×1
EMNLP
×1
SIGKDD
×1
Frequent coauthors
Shiyao Cui
DBLP profile ↗
ORCID search ↗
×3
Junxiao Yang
DBLP profile ↗
ORCID search ↗
×2
Renmiao Chen
DBLP profile ↗
ORCID search ↗
×1
Shangqing Tu
DBLP profile ↗
ORCID search ↗
×1
Yida Lu
DBLP profile ↗
ORCID search ↗
×1
Papers
How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study
ACL 2026
·
Zhexin Zhang
LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety
ACL 2026
·
Junxiao Yang
DBLP profile ↗
ORCID search ↗
New Terms, New Toxicity: Consensus-based Chinese Neologism Toxicity Detection via Search-Augmented LLMs
ACL 2026
·
Shiyao Cui
DBLP profile ↗
ORCID search ↗
When Smiley Turns Hostile: Interpreting How Emojis Trigger LLMs' Toxicity
AAAI 2026
·
Shiyao Cui
DBLP profile ↗
ORCID search ↗
Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints
ACL 2025
·
Junxiao Yang
DBLP profile ↗
ORCID search ↗
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
ACM MM 2025
·
Renmiao Chen
DBLP profile ↗
ORCID search ↗
Knowledge-to-Jailbreak: Investigating Knowledge-driven Jailbreaking Attacks for Large Language Models
SIGKDD 2025
·
Shangqing Tu
DBLP profile ↗
ORCID search ↗
LongSafety: Evaluating Long-Context Safety of Large Language Models
ACL 2025
·
Yida Lu
DBLP profile ↗
ORCID search ↗
ShieldVLM: Safeguarding the Multimodal Implicit Toxicity via Deliberative Reasoning with LVLMs: ShieldVLM
ACM MM 2025
·
Shiyao Cui
DBLP profile ↗
ORCID search ↗
A Design of Interface for Visual-Impaired People to Access Visual Information from Images Featuring Large Language Models and Visual Language Models
CHI 2024
·
Zhexin Zhang
Defending Large Language Models Against Jailbreaking Attacks Through Goal Prioritization
ACL 2024
·
Zhexin Zhang
SafetyBench: Evaluating the Safety of Large Language Models
ACL 2024
·
Zhexin Zhang
ShieldLM: Empowering LLMs as Aligned, Customizable and Explainable Safety Detectors
EMNLP 2024
·
Zhexin Zhang