P
PaperPicks
Conferences
Yudong Chen
12 papers at tracked venues · 9 at CORE A* · active 2024–2025
DBLP profile ↗
ORCID search ↗
Venues
ICML
×4
NeurIPS
×4
AISTATS
×2
AAAI
×1
COLT
×1
Frequent coauthors
Brahma S. Pavse
DBLP profile ↗
ORCID search ↗
×2
Dongyan Lucy Huo
DBLP profile ↗
ORCID search ↗
×2
Matthew Zurek
DBLP profile ↗
ORCID search ↗
×2
Yuanhe Zhang
DBLP profile ↗
ORCID search ↗
×1
Yichen Wang
DBLP profile ↗
ORCID search ↗
×1
Jeongyeol Kwon
DBLP profile ↗
ORCID search ↗
×1
Young Wu
DBLP profile ↗
ORCID search ↗
×1
Emmanouil-Vasileios Vlatakis-Gkaragkounis
DBLP profile ↗
ORCID search ↗
×1
Tyler Sam
DBLP profile ↗
ORCID search ↗
×1
Papers
LoRA-One: One-Step Full Gradient Could Suffice for Fine-Tuning Large Language Models, Provably and Efficiently
ICML 2025
·
Yuanhe Zhang
DBLP profile ↗
ORCID search ↗
Stable Offline Value Function Learning with Bisimulation-based Representations
ICML 2025
·
Brahma S. Pavse
DBLP profile ↗
ORCID search ↗
The φ Curve: The Shape of Generalization through the Lens of Norm-based Capacity Control
NeurIPS 2025
·
Yichen Wang
DBLP profile ↗
ORCID search ↗
Two-Timescale Linear Stochastic Approximation: Constant Stepsizes Go a Long Way
AISTATS 2025
·
Jeongyeol Kwon
DBLP profile ↗
ORCID search ↗
Effectiveness of Constant Stepsize in Markovian LSA and Statistical Inference
AAAI 2024
·
Dongyan Lucy Huo
DBLP profile ↗
ORCID search ↗
Gap-Free Clustering: Sensitivity and Robustness of SDP
COLT 2024
·
Matthew Zurek
DBLP profile ↗
ORCID search ↗
Learning to Stabilize Online Reinforcement Learning in Unbounded State Spaces
ICML 2024
·
Brahma S. Pavse
DBLP profile ↗
ORCID search ↗
Minimally Modifying a Markov Game to Achieve Any Nash Equilibrium and Value
ICML 2024
·
Young Wu
DBLP profile ↗
ORCID search ↗
Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward MDPs
NeurIPS 2024
·
Matthew Zurek
DBLP profile ↗
ORCID search ↗
Stochastic Methods in Variational Inequalities: Ergodicity, Bias and Refinements
AISTATS 2024
·
Emmanouil-Vasileios Vlatakis-Gkaragkounis
DBLP profile ↗
ORCID search ↗
The Collusion of Memory and Nonlinearity in Stochastic Approximation With Constant Stepsize
NeurIPS 2024
·
Dongyan Lucy Huo
DBLP profile ↗
ORCID search ↗
The Limits of Transfer Reinforcement Learning with Latent Low-rank Structure
NeurIPS 2024
·
Tyler Sam
DBLP profile ↗
ORCID search ↗