PPaperPicks

Yixu Wang

15 papers at tracked venues · 12 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models
    ACL 2026 · Yixu Wang
  2. Probing the Safety Robustness of LLMs in Latent Space
  3. The Other Mind: How Language Models Exhibit Human Temporal Cognition
  4. A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos
  5. Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes?
  6. HoneypotNet: Backdoor Attacks Against Model Extraction
    AAAI 2025 · Yixu Wang
  7. Ideator: Jailbreaking and Benchmarking Large Vision-Language Models Using Themselves
  8. JailBound: Jailbreaking Internal Safety Boundaries of Vision-Language Models
  9. Reflection-Bench: Evaluating Epistemic Agency in Large Language Models
  10. SafeVid: Toward Safety Aligned Video Large Multimodal Models
    NeurIPS 2025 · Yixu Wang
  11. StolenLoRA: Exploring LoRA Extraction Attacks via Synthetic Data
    ICCV 2025 · Yixu Wang
  12. ESC-Eval: Evaluating Emotion Support Conversations in Large Language Models
  13. Fake Alignment: Are LLMs Really Aligned Well?
    NAACL 2024 · Yixu Wang
  14. Flames: Benchmarking Value Alignment of LLMs in Chinese
  15. MLLMGuard: A Multi-dimensional Safety Evaluation Suite for Multimodal Large Language Models