P
PaperPicks
Conferences
Zishun Yu
5 papers at tracked venues · 4 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
AAAI
×2
ICLR
×1
ICML
×1
UAI
×1
Frequent coauthors
Wenzhe Fan
DBLP profile ↗
ORCID search ↗
×1
Papers
Language Model Distillation: A Temporal Difference Imitation Learning Perspective
AAAI 2026
·
Zishun Yu
Think Smarter not Harder: Adaptive Reasoning with Inference Aware Optimization
ICML 2025
·
Zishun Yu
Towards Efficient Collaboration via Graph Modeling in Reinforcement Learning
AAAI 2025
·
Wenzhe Fan
DBLP profile ↗
ORCID search ↗
B-Coder: Value-Based Deep Reinforcement Learning for Program Synthesis
ICLR 2024
·
Zishun Yu
Offline Reward Perturbation Boosts Distributional Shift in Online RL
UAI 2024
·
Zishun Yu