PPaperPicks

Yingchun Wang

Shanghai Artificial Intelligence Laboratory, China

19 papers at tracked venues · 15 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models
  2. Deliberative Searcher: Improving LLM Reliability via Reinforcement Learning with Constraints
  3. Probing the Safety Robustness of LLMs in Latent Space
  4. The Other Mind: How Language Models Exhibit Human Temporal Cognition
  5. A Mousetrap: Fooling Large Reasoning Models for Jailbreak with Chain of Iterative Chaos
  6. Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes?
  7. Beyond Correctness: Confidence-Aware Reward Modeling for Enhancing Large Language Model Reasoning
  8. From Evasion to Concealment: Stealthy Knowledge Unlearning for LLMs
  9. HoneypotNet: Backdoor Attacks Against Model Extraction
  10. IDMR: Towards Instance-Driven Precise Visual Correspondence in Multimodal Retrieval
  11. Ideator: Jailbreaking and Benchmarking Large Vision-Language Models Using Themselves
  12. JailBound: Jailbreaking Internal Safety Boundaries of Vision-Language Models
  13. Reflection-Bench: Evaluating Epistemic Agency in Large Language Models
  14. SafeVid: Toward Safety Aligned Video Large Multimodal Models
  15. StolenLoRA: Exploring LoRA Extraction Attacks via Synthetic Data
  16. ESC-Eval: Evaluating Emotion Support Conversations in Large Language Models
  17. Fake Alignment: Are LLMs Really Aligned Well?
  18. Flames: Benchmarking Value Alignment of LLMs in Chinese
  19. MLLMGuard: A Multi-dimensional Safety Evaluation Suite for Multimodal Large Language Models