PPaperPicks

Fengshuo Bai

9 papers at tracked venues · 6 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. ToolPRM: Fine-Grained Inference Scaling of Structured Outputs for Function Calling
  2. AdaptFlow: Adaptive Workflow Optimization via Meta-Learning
  3. Amulet: ReAlignment During Test Time for Personalized Preference Adaptation of LLMs
  4. DexFlyWheel: A Scalable and Self-improving Data Generation Framework for Dexterous Manipulation
  5. GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation
  6. RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors
    AAAI 2025 · Fengshuo Bai
  7. STAR: Efficient Preference-based Reinforcement Learning via Dual Regularization
    NeurIPS 2025 · Fengshuo Bai
  8. β-DQN: Improving Deep Q-Learning By Evolving the Behavior
  9. PEARL: Zero-shot Cross-task Preference Alignment and Robust Reward Learning for Robotic Manipulation