PPaperPicks

Tangjie Lv

23 papers at tracked venues · 16 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Ψ-Arena: Interactive Assessment and Optimization of LLM-based Psychological Counselors with Tripartite Feedback
  2. CharacterBench: Benchmarking Character Customization of Large Language Models
  3. Crisp: Cognitive Restructuring of Negative Thoughts through Multi-turn Supportive Dialogues
  4. Empowering Economic Simulation for Massively Multiplayer Online Games through Generative Agent-Based Modeling
  5. Reinforcement Learning from Imperfect Corrective Actions and Proxy Rewards
  6. StoryWeaver: A Unified World Model for Knowledge-Enhanced Story Character Customization
  7. Towards Transferable Personality Representation Learning based on Triplet Comparisons and Its Applications
  8. A Trajectory Perspective on the Role of Data Sampling Techniques in Offline Reinforcement Learning
  9. AlignDiff: Aligning Diverse Human Preferences via Behavior-Customisable Diffusion Model
  10. Bayesian Design Principles for Offline-to-Online Reinforcement Learning
  11. EnMatch: Matchmaking for Better Player Engagement via Neural Combinatorial Optimization
  12. Hybrid CtrlFormer: Learning Adaptive Search Space Partition for Hybrid Action Control via Transformer-based Monte Carlo Tree Search
  13. MGMatch: Fast Matchmaking with Nonlinear Objective and Constraints via Multimodal Deep Graph Learning
  14. Optimistic Value Instructors for Cooperative Multi-Agent Reinforcement Learning
  15. STAR: Spatio-Temporal State Compression for Multi-Agent Tasks with Rich Observations
  16. Structure-CLIP: Towards Scene Graph Knowledge to Enhance Multi-Modal Structured Representations
  17. Stylized Offline Reinforcement Learning: Extracting Diverse High-Quality Behaviors from Heterogeneous Datasets
  18. Temporal Uplift Modeling for Online Marketing
  19. The MMO Economist: AI Empowers Robust, Healthy, and Sustainable P2W MMO Economies
  20. Towards a Simultaneous and Granular Identity-Expression Control in Personalized Face Generation
  21. Unlock the Intermittent Control Ability of Model Free Reinforcement Learning
  22. vMFER: Von Mises-Fisher Experience Resampling Based on Uncertainty of Gradient Directions for Policy Improvement
  23. vMFER: von Mises-Fisher Experience Resampling Based on Uncertainty of Gradient Directions for Policy Improvement of Actor-Critic Algorithms