PPaperPicks

Junxiao Yang

5 papers at tracked venues · 5 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study
  2. LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety
    ACL 2026 · Junxiao Yang
  3. When Smiley Turns Hostile: Interpreting How Emojis Trigger LLMs' Toxicity
  4. Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints
    ACL 2025 · Junxiao Yang
  5. Defending Large Language Models Against Jailbreaking Attacks Through Goal Prioritization