PPaperPicks

Haodong Zhao

7 papers at tracked venues · 6 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Turning Failures into Value: Negative Experience Replay for RLVR via Confidence Gating and Boundary Failure Sampling
  2. VANE: Guiding High-Value Exploration in RLVR via Outcome-Process Novelty Shaping
  3. AMoPO: Adaptive Multi-objective Preference Optimization without Reward Models and Reference Models
  4. Watch Out Your Album! On the Inadvertent Privacy Memorization in Multi-Modal Large Language Models
  5. When to Continue Thinking: Adaptive Thinking Mode Switching for Efficient Reasoning
  6. Revisiting the Information Capacity of Neural Network Watermarks: Upper Bound Estimation and Beyond
  7. UOR: Universal Backdoor Attacks on Pre-trained Language Models