PPaperPicks

Shunyu Liu

Nanyang Technological University, Singapore

30 papers at tracked venues · 23 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Dual-branch Spatial-Temporal Self-supervised Representation for Enhanced Road Network Learning
  2. Reasoning-Guided Exploration for Online DPO
  3. RubricHub: A Comprehensive and Highly Discriminative Rubric Dataset via Automated Coarse-to-Fine Generation
  4. Agent-Aware Training for Agent-Agnostic Action Advising in Deep Reinforcement Learning
  5. Bi-Level Mean Field: Dynamic Grouping for Large-Scale MARL
  6. CADP: Towards Better Centralized Learning for Decentralized Execution in MARL
  7. CADP: Towards Better Centralized Learning for Decentralized Execution in MARL
  8. Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning
  9. Cooperative Policy Agreement: Learning Diverse Policy for Offline MARL
  10. Disentangled Condensation for Large-scale Graphs
  11. Disentangled Table-Graph Representation for Interpretable Transmission Line Fault Location
  12. Dynamic Parallel Tree Search for Efficient LLM Reasoning
  13. From GNNs to Trees: Multi-Granular Interpretability for Graph Neural Networks
  14. Holistic Semantic Representation for Navigational Trajectory Generation
  15. Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search
  16. Odyssey : Empowering Minecraft Agents with Open-World Skills
    IJCAI 2025 · Shunyu Liu
  17. Powerformer: A Section-adaptive Transformer for Power Flow Adjustment
  18. R1-VL: Learning to Reason with Multimodal Large Language Models via Step-Wise Group Relative Policy Optimization
  19. SPAZER: Spatial-Semantic Progressive Reasoning Agent for Zero-shot 3D Visual Grounding
  20. SeRL: Self-play Reinforcement Learning for Large Language Models with Limited Data
  21. Supervised Optimism Correction: Be Confident When LLMs Are Sure
  22. Tree of Preferences for Diversified Recommendation
  23. VORTA: Efficient Video Diffusion via Routing Sparse Attention
  24. A2PO: Towards Effective Offline Reinforcement Learning from an Advantage-aware Perspective
  25. COLA: Cross-city Mobility Transformer for Human Trajectory Simulation
  26. Improving Adversarial Robustness via Feature Pattern Consistency Constraint
  27. Learning a Mini-Batch Graph Transformer via Two-Stage Interaction Augmentation
  28. Simple Graph Condensation
  29. Temporal Prototype-Aware Learning for Active Voltage Control on Power Distribution Networks
  30. Unveiling Global Interactive Patterns across Graphs: Towards Interpretable Graph Neural Networks