PPaperPicks

Bo Ding

13 papers at tracked venues · 11 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Divergence or Convergence? A Deep Insight into the Crowd Collaboration and its Productivity in Open Source Software based on Entropy
  2. EvoNarrator: Modeling Scientific Evolution for Feasible Hypothesis Generation
  3. Empowering Large Language Model Agent through Step-Level Self-Critique and Self-Training
  4. Enhancing Decision-Making for LLM Agents via Step-Level Q-Value Models
  5. Improving the Continuity of Goal-Achievement Ability via Policy Self-Regularization for Goal-Conditioned Reinforcement Learning
  6. Preference-Strength-Aware Self-Improving Alignment with Generative Preference Models
  7. V-Pilot: A Velocity Vector Control Agent for Fixed-Wing UAVs from Imperfect Demonstrations
  8. VVC-Gym: A Fixed-Wing UAV Reinforcement Learning Environment for Multi-Goal Long-Horizon Problems
  9. Goal-Conditioned On-Policy Reinforcement Learning
  10. Iterative Regularized Policy Optimization with Imperfect Demonstrations
  11. Optimistic Model Rollouts for Pessimistic Offline Policy Optimization
  12. Selective Learning for Sample-Efficient Training in Multi-Agent Sparse Reward Tasks (Extended Abstract)
  13. Tracing Training Progress: Dynamic Influence Based Selection for Active Learning