P
PaperPicks
Conferences
Fengshuo Bai
9 papers at tracked venues · 6 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
NeurIPS
×2
AAAI
×1
AAMAS
×1
ACL
×1
EMNLP
×1
ICLR
×1
ICML
×1
NAACL
×1
Frequent coauthors
Runchuan Zhu
DBLP profile ↗
ORCID search ↗
×2
Jianghao Lin
DBLP profile ↗
ORCID search ↗
×1
Zhaowei Zhang
DBLP profile ↗
ORCID search ↗
×1
Kefei Zhu
DBLP profile ↗
ORCID search ↗
×1
Hongming Zhang
DBLP profile ↗
ORCID search ↗
×1
Runze Liu
DBLP profile ↗
ORCID search ↗
×1
Papers
ToolPRM: Fine-Grained Inference Scaling of Structured Outputs for Function Calling
ACL 2026
·
Jianghao Lin
DBLP profile ↗
ORCID search ↗
AdaptFlow: Adaptive Workflow Optimization via Meta-Learning
EMNLP 2025
·
Runchuan Zhu
DBLP profile ↗
ORCID search ↗
Amulet: ReAlignment During Test Time for Personalized Preference Adaptation of LLMs
ICLR 2025
·
Zhaowei Zhang
DBLP profile ↗
ORCID search ↗
DexFlyWheel: A Scalable and Self-improving Data Generation Framework for Dexterous Manipulation
NeurIPS 2025
·
Kefei Zhu
DBLP profile ↗
ORCID search ↗
GRAIT: Gradient-Driven Refusal-Aware Instruction Tuning for Effective Hallucination Mitigation
NAACL 2025
·
Runchuan Zhu
DBLP profile ↗
ORCID search ↗
RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors
AAAI 2025
·
Fengshuo Bai
STAR: Efficient Preference-based Reinforcement Learning via Dual Regularization
NeurIPS 2025
·
Fengshuo Bai
β-DQN: Improving Deep Q-Learning By Evolving the Behavior
AAMAS 2025
·
Hongming Zhang
DBLP profile ↗
ORCID search ↗
PEARL: Zero-shot Cross-task Preference Alignment and Robust Reward Learning for Robotic Manipulation
ICML 2024
·
Runze Liu
DBLP profile ↗
ORCID search ↗