PPaperPicks

Xiaojie Wang

8 papers at tracked venues · 6 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Enhancing Multilingual Reasoning via Steerable Model Merging
  2. Multi-Agent-as-Judge: Aligning LLM-Agent-Based Automated Evaluation with Multi-Dimensional Human Evaluation
  3. CoTD-PO: Chain-of-Thought Distillation with Preference Optimization
  4. Collab-Overcooked: Benchmarking and Evaluating Large Language Models as Collaborative Agents
  5. Controlled Low-Rank Adaptation with Subspace Regularization for Continued Training on Large Language Models
  6. Data with High and Consistent Preference Difference Are Better for Reward Model
  7. Non-asymptotic Error Bounds in W2-Distance with Sqrt(d) Dimension Dependence and First Order Convergence for Langevin Monte Carlo beyond Log-Concavity
  8. Phased Instruction Fine-Tuning for Large Language Models