PPaperPicks

Jiancan Wu

26 papers at tracked venues · 22 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Delayed Feedback Modeling with Influence Functions
  2. SepSeq: A Training-Free Framework for Long Numerical Sequence Processing in LLMs
  3. Task-Aligned Unlearning for Multimodal Large Language Models
  4. Addressing Missing Data Issue for Diffusion-based Recommendation
  5. AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization
  6. Fading to Grow: Growing Preference Ratios via Preference Fading Discrete Diffusion for Recommendation
  7. LANCE: Exploration and Reflection for LLM-based Textual Attacks on News Recommender Systems
  8. LaMP-Val: Large Language Models Empower Personalized Valuation in Auction
  9. Larger or Smaller Reward Margins to Select Preferences for LLM Alignment?
  10. On Efficiency-Effectiveness Trade-off of Diffusion-based Recommenders
  11. Personal Travel Solver: A Preference-Driven LLM-Solver System for Travel Planning
  12. RePO: Understanding Preference Learning Through ReLU-Based Optimization
  13. Robust Preference Optimization via Dynamic Target Margins
  14. Think before Recommendation: Autonomous Reasoning-enhanced Recommender
  15. Towards Large Generative Recommendation: A Tokenization Perspective
  16. Towards Robust Alignment of Language Models: Distributionally Robustifying Direct Preference Optimization
  17. Unified Parameter-Efficient Unlearning for LLMs
  18. BSL: Understanding and Improving Softmax Loss for Recommendation
  19. Customizing Language Models with Instance-wise LoRA for Sequential Recommendation
  20. Dynamic Sparse Learning: A Novel Paradigm for Efficient Recommendation
  21. LLaRA: Large Language-Recommendation Assistant
  22. Leave No Patient Behind: Enhancing Medication Recommendation for Rare Disease Patients
  23. Let Me Do It For You: Towards LLM Empowered Recommendation via Tool Learning
  24. Masked Graph Modeling with Multi- View Contrast
  25. MuggleMath: Assessing the Impact of Query and Response Augmentation on Math Reasoning
  26. β-DPO: Direct Preference Optimization with Dynamic β