PPaperPicks

Junjie Ye

Fudan University, School of Computer Science, Shanghai, China

22 papers at tracked venues · 16 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AgentGym2: Benchmarking Large Language Model Agents in De-Idealized Real-World Environments
  2. Beyond Scaling: Measuring and Predicting the Upper Bound of Knowledge Retention in Language Model Pre-Training
  3. DARM: Distribution-Aware Reward Modeling by Alleviating Biases from Low Preference-Context Dependency Data
  4. Feedback-Driven Tool-Use Improvements in Large Language Models via Automated Build Environments
    ACL 2026 · Junjie Ye
  5. FinToolSyn: A forward synthesis Framework for Financial Tool-Use Dialogue Data with Dynamic Tool Retrieval
  6. MetaAct-RL: Training Language Models for Reasoning Through Meta-Action-Based Reinforcement Learning
  7. MulDimIF: A Multi-Dimensional Constraint Framework for Evaluating and Improving Instruction Following in Large Language Models
    ACL 2026 · Junjie Ye
  8. Muse: Towards Reproducible Long-Form Song Generation with Fine-Grained Style Control
  9. VRPO: Rethinking Value Modeling for Robust RL under Noisy Supervision in LLM Post-Training
  10. What Makes a Good Speech Tokenizer for LLM-Centric Speech Generation? A Systematic Study
  11. Alleviating Shifted Distribution in Human Preference Alignment through Meta-Learning
  12. Analyzing the Effects of Supervised Fine-Tuning on Model Knowledge from Token and Parameter Levels
    EMNLP 2025 · Junjie Ye
  13. CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios
  14. Measuring Data Diversity for Instruction Tuning: A Systematic Analysis and A Reliable Metric
  15. TL-Training: A Task-Feature-Based Framework for Training Large Language Models in Tool Use
    EMNLP 2025 · Junjie Ye
  16. ToolHop: A Query-Driven Benchmark for Evaluating Large Language Models in Multi-Hop Tool Use
    ACL 2025 · Junjie Ye
  17. Improving Discriminative Capability of Reward Models in RLHF Using Contrastive Learning
  18. LLM can Achieve Self-Regulation via Hyperparameter Aware Generation
  19. Linear Alignment: A Closed-form Solution for Aligning Human Preferences without Tuning and Feedback
  20. RoTBench: A Multi-Level Benchmark for Evaluating the Robustness of Large Language Models in Tool Learning
    EMNLP 2024 · Junjie Ye
  21. ToolSword: Unveiling Safety Issues of Large Language Models in Tool Learning Across Three Stages
    ACL 2024 · Junjie Ye
  22. TransferTOD: A Generalizable Chinese Multi-Domain Task-Oriented Dialogue System with Transfer Capabilities