PPaperPicks

Pradeep Varakantham

21 papers at tracked venues · 15 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Optimizing Ride-Pooling Operations with Extended Pickup and Drop-Off Flexibility
  2. Bootstrapping Language Models with DPO Implicit Rewards
  3. EduQate: Generating Adaptive Curricula through RMABs in Education Settings
  4. Marginal Benefit Driven RL Teacher for Unsupervised Environment Design
  5. No Experts, No Problem: Avoidance Learning from Bad Demonstrations
  6. Offline Safe Reinforcement Learning Using Trajectory Classification
  7. On Generalization Across Environments In Multi-Objective Reinforcement Learning
  8. On Learning Informative Trajectory Embeddings for Imitation, Classification and Regression
  9. On Minimizing Adversarial Counterfactual Error in Adversarial Reinforcement Learning
  10. Semantic Loss Guided Data Efficient Supervised Fine Tuning for Safe Responses in LLMs
  11. Unlocking the Planning Capabilities of Large Language Models with Maximum Diversity Fine-tuning
  12. Handling Long and Richly Constrained Tasks through Constrained Hierarchical Reinforcement Learning
  13. Imitate the Good and Avoid the Bad: An Incremental Approach to Safe Reinforcement Learning
  14. Improving Environment Novelty Quantification for Effective Unsupervised Environment Design
  15. Preserving the Privacy of Reward Functions in MDPs through Deception
  16. Regret-based Defense in Adversarial Reinforcement Learning
  17. Reward Penalties on Augmented States for Solving Richly Constrained RL Effectively
  18. SPRINQL: Sub-optimal Demonstrations driven Offline Imitation Learning
  19. Safety through feedback in Constrained RL
  20. Unifying Regret and State-Action Space Coverage for Effective Unsupervised Environment Design
  21. Unsupervised Training Sequence Design: Efficient and Generalizable Agent Training