PPaperPicks

Xiangzheng Zhang

12 papers at tracked venues · 8 at CORE A* · active 20252026

Venues

Frequent coauthors

Papers

  1. DMN: A Compositional Framework for Jailbreaking Multimodal LLMs with Multi-Image Inputs
  2. Efficient Switchable Safety Control in LLMs via Magic-Token-Guided Co-Training
  3. Light-IF: Endowing LLMs with Generalizable Reasoning via Preview and Self-Checking for Complex Instruction Following
  4. Thinking with Reasoning Skills: Fewer Tokens, More Accuracy
  5. TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment
  6. When Good OCR Is Not Enough: Benchmarking OCR Robustness for Retrieval-Augmented Generation
  7. Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision
  8. Expand VSR Benchmark for VLLM to Expertize in Spatial Rules
  9. Large Language Models Badly Generalize across Option Length, Problem Types, and Irrelevant Noun Replacements
  10. Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond
  11. Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models
  12. Router Upcycling: Leveraging Mixture-of-Routers in Mixture-of-Experts Upcycling