PPaperPicks

Yudong Chen

12 papers at tracked venues · 9 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. LoRA-One: One-Step Full Gradient Could Suffice for Fine-Tuning Large Language Models, Provably and Efficiently
  2. Stable Offline Value Function Learning with Bisimulation-based Representations
  3. The φ Curve: The Shape of Generalization through the Lens of Norm-based Capacity Control
  4. Two-Timescale Linear Stochastic Approximation: Constant Stepsizes Go a Long Way
  5. Effectiveness of Constant Stepsize in Markovian LSA and Statistical Inference
  6. Gap-Free Clustering: Sensitivity and Robustness of SDP
  7. Learning to Stabilize Online Reinforcement Learning in Unbounded State Spaces
  8. Minimally Modifying a Markov Game to Achieve Any Nash Equilibrium and Value
  9. Span-Based Optimal Sample Complexity for Weakly Communicating and General Average Reward MDPs
  10. Stochastic Methods in Variational Inequalities: Ergodicity, Bias and Refinements
    AISTATS 2024 ·
    Emmanouil-Vasileios Vlatakis-Gkaragkounis
  11. The Collusion of Memory and Nonlinearity in Stochastic Approximation With Constant Stepsize
  12. The Limits of Transfer Reinforcement Learning with Latent Low-rank Structure