PPaperPicks

Mengdi Wang

Princeton University, Center for Statistics and Machine Learning, Department of Electrical and Computer Engineering, NJ, USA

44 papers at tracked venues · 38 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Jailbreaks as Inference-Time Alignment: A Framework for Understanding Safety Failures in LLMs
  2. A Common Pitfall of Margin-based Language Model Alignment: Gradient Entanglement
  3. A First-order Generative Bilevel Optimization Framework for Diffusion Models
  4. Co-Evolving LLM Coder and Unit Tester via Reinforcement Learning
  5. Collab: Controlled Decoding using Mixture of Agents for LLM Alignment
  6. DISC: Dynamic Decomposition Improves LLM Inference Scaling
  7. Deep Reinforcement Learning for Efficient and Fair Allocation of Healthcare Resources
  8. Diffusion Transformer Captures Spatial-Temporal Dependencies: A Theory for Gaussian Process Data
  9. Does Thinking More Always Help? Mirage of Test-Time Scaling in Reasoning Models
  10. Emergent Symbolic Mechanisms Support Abstract Reasoning in Large Language Models
  11. EmoAgent: Assessing and Safeguarding Human-AI Interaction for Mental Health Safety
  12. Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment
  13. IterComp: Iterative Composition-Aware Feedback Learning from Model Gallery for Text-to-Image Generation
  14. MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
  15. MMaDA: Multimodal Large Diffusion Language Models
  16. Preacher: Paper-to-Video Agentic System
  17. ReasonFlux-PRM: Trajectory-Aware PRMs for Long Chain-of-Thought Reasoning in LLMs
  18. Rectified Diffusion: Straightness Is Not Your Need in Rectified Flow
  19. Securing the Language of Life: Inheritable Watermarks from DNA Language Models to Proteins
  20. Temporal Consistency for LLM Reasoning Process Error Identification
  21. Towards Understanding Text Hallucination of Diffusion Models via Local Generation Bias
  22. Training-Free Guidance Beyond Differentiability: Scalable Path Steering with Tree Search in Diffusion and Flow Models
  23. TreeBoN: Enhancing Inference-Time Alignment with Speculative Tree-Search and Best-of-N Sampling
  24. A Theoretical Perspective for Speculative Decoding Algorithm
  25. Assessing the Brittleness of Safety Alignment via Pruning and Low-Rank Modifications
  26. Conversational Dueling Bandits in Generalized Linear Models
  27. Fast Best-of-N Decoding via Speculative Rejection
  28. FlexSBDD: Structure-Based Drug Design with Flexible Protein Modeling
  29. Global Convergence in Training Large-Scale Transformers
  30. Gradient Guidance for Diffusion Models: An Optimization Perspective
  31. Information-Directed Pessimism for Offline Reinforcement Learning
  32. Is Inverse Reinforcement Learning Harder than Standard Reinforcement Learning? A Theoretical Perspective
  33. MaxMin-RLHF: Alignment with Diverse Human Preferences
  34. Nonparametric Classification on Low Dimensional Manifolds using Overparameterized Convolutional Residual Networks
  35. Offline Multitask Representation Learning for Reinforcement Learning
  36. One-Layer Transformer Provably Learns One-Nearest Neighbor In Context
  37. PARL: A Unified Framework for Policy Alignment in Reinforcement Learning from Human Feedback
  38. Policy Evaluation for Reinforcement Learning from Human Feedback: A Sample Complexity Analysis
  39. Sample-Efficient Learning of POMDPs with Multiple Observations In Hindsight
  40. Theoretical insights for diffusion guidance: A case study for Gaussian mixture models
  41. Theory of Consistency Diffusion Models: Distribution Estimation Meets Fast Sampling
  42. Transfer Q-star : Principled Decoding for LLM Alignment
  43. Tree Search-Based Evolutionary Bandits for Protein Sequence Optimization
  44. Visual Adversarial Examples Jailbreak Aligned Large Language Models