P
PaperPicks
Conferences
Yasheng Wang
34 papers at tracked venues · 25 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
ACL
×12
ICLR
×6
EMNLP
×5
NeurIPS
×4
AAAI
×2
NAACL
×2
CIKM
×1
ECML-PKDD
×1
ICML
×1
Frequent coauthors
Zezhong Wang
DBLP profile ↗
ORCID search ↗
×3
Lingyue Fu
DBLP profile ↗
ORCID search ↗
×2
Yuxin Jiang
DBLP profile ↗
ORCID search ↗
×2
Xiangyang Li
DBLP profile ↗
ORCID search ↗
×2
Qingyao Li
DBLP profile ↗
ORCID search ↗
×2
Qiyuan Zhang
DBLP profile ↗
ORCID search ↗
×2
Fan Gao
DBLP profile ↗
ORCID search ↗
×1
Xingshan Zeng
DBLP profile ↗
ORCID search ↗
×1
Wenjun Li
DBLP profile ↗
ORCID search ↗
×1
Kuicai Dong
DBLP profile ↗
ORCID search ↗
×1
Kounianhua Du
DBLP profile ↗
ORCID search ↗
×1
Jizheng Chen
DBLP profile ↗
ORCID search ↗
×1
Papers
EssayBench: Evaluating Large Language Models in Multi-Genre Chinese Essay Writing
AAAI 2026
·
Fan Gao
DBLP profile ↗
ORCID search ↗
ToolACE-R: Model-aware Iterative Training and Adaptive Refinement for Tool learning
AAAI 2026
·
Xingshan Zeng
DBLP profile ↗
ORCID search ↗
Adaptive Tool Use in Large Language Models with Meta-Cognition Trigger
ACL 2025
·
Wenjun Li
DBLP profile ↗
ORCID search ↗
AdvKT: An Adversarial Multi-step Training Framework for Knowledge Tracing
ECML-PKDD 2025
·
Lingyue Fu
DBLP profile ↗
ORCID search ↗
Benchmarking Retrieval-Augmented Multimomal Generation for Document Question Answering
NeurIPS 2025
·
Kuicai Dong
DBLP profile ↗
ORCID search ↗
Boost, Disentangle, and Customize: A Robust System2-to-System1 Pipeline for Code Generation
ACL 2025
·
Kounianhua Du
DBLP profile ↗
ORCID search ↗
Bridging and Modeling Correlations in Pairwise Data for Direct Preference Optimization
ICLR 2025
·
Yuxin Jiang
DBLP profile ↗
ORCID search ↗
Chain-of-Probe: Examining the Necessity and Accuracy of CoT Step-by-Step
NAACL 2025
·
Zezhong Wang
DBLP profile ↗
ORCID search ↗
CoIR: A Comprehensive Benchmark for Code Information Retrieval Models
ACL 2025
·
Xiangyang Li
DBLP profile ↗
ORCID search ↗
CodePRM: Execution Feedback-enhanced Process Reward Model for Code Generation
ACL 2025
·
Qingyao Li
DBLP profile ↗
ORCID search ↗
Crowd Comparative Reasoning: Unlocking Comprehensive Evaluations for LLM-as-a-Judge
ACL 2025
·
Qiyuan Zhang
DBLP profile ↗
ORCID search ↗
DebateCoder: Towards Collective Intelligence of LLMs via Test Case Driven LLM Debate for Code Generation
ACL 2025
·
Jizheng Chen
DBLP profile ↗
ORCID search ↗
DeepDiver: Adaptive Web-Search Intensity Scaling via Reinforcement Learning
NeurIPS 2025
·
Wenxuan Shi
DBLP profile ↗
ORCID search ↗
Flat-LoRA: Low-Rank Adaptation over a Flat Loss Landscape
ICML 2025
·
Tao Li
DBLP profile ↗
ORCID search ↗
Humanity's Last Code Exam: Can Advanced LLMs Conquer Human's Hardest Code Competition?
EMNLP 2025
·
Xiangyang Li
DBLP profile ↗
ORCID search ↗
Instruction-Tuning Data Synthesis from Scratch via Web Reconstruction
ACL 2025
·
Yuxin Jiang
DBLP profile ↗
ORCID search ↗
Learning Evolving Tools for Large Language Models
ICLR 2025
·
Guoxin Chen
DBLP profile ↗
ORCID search ↗
NILE: Internal Consistency Alignment in Large Language Models
EMNLP 2025
·
Minda Hu
DBLP profile ↗
ORCID search ↗
NL-Debugging: Exploiting Natural Language as an Intermediate Representation for Code Debugging
EMNLP 2025
·
Weiming Zhang
DBLP profile ↗
ORCID search ↗
Proactive Agent: Shifting LLM Agents from Reactive Responses to Active Assistance
ICLR 2025
·
Yaxi Lu
DBLP profile ↗
ORCID search ↗
QFFT, Question-Free Fine-Tuning for Adaptive Reasoning
NeurIPS 2025
·
Wanlong Liu
DBLP profile ↗
ORCID search ↗
RethinkMCTS: Refining Erroneous Thoughts in Monte Carlo Tree Search for Code Generation
EMNLP 2025
·
Qingyao Li
DBLP profile ↗
ORCID search ↗
RevisEval: Improving LLM-as-a-Judge via Response-Adapted References
ICLR 2025
·
Qiyuan Zhang
DBLP profile ↗
ORCID search ↗
RidgeLoRA: Matrix Ridge Enhanced Low-Rank Adaptation of Large Language Models
NeurIPS 2025
·
Junda Zhu
DBLP profile ↗
ORCID search ↗
Safe: Enhancing Mathematical Reasoning in Large Language Models via Retrospective Step-aware Formal Verification
ACL 2025
·
Chengwu Liu
DBLP profile ↗
ORCID search ↗
Spa-Bench: a comprehensive Benchmark for Smartphone Agent Evaluation
ICLR 2025
·
Jingxuan Chen
DBLP profile ↗
ORCID search ↗
Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning
EMNLP 2025
·
Zezhong Wang
DBLP profile ↗
ORCID search ↗
ToolACE: Winning the Points of LLM Function Calling
ICLR 2025
·
Weiwen Liu
DBLP profile ↗
ORCID search ↗
ToolFlow: Boosting LLM Tool-Calling Through Natural and Coherent Dialogue Synthesis
NAACL 2025
·
Zezhong Wang
DBLP profile ↗
ORCID search ↗
Dynamic Stochastic Decoding Strategy for Open-Domain Dialogue Generation
ACL 2024
·
Yiwei Li
DBLP profile ↗
ORCID search ↗
Evaluating Robustness of Generative Search Engine on Adversarial Factoid Questions
ACL 2024
·
Xuming Hu
DBLP profile ↗
ORCID search ↗
Planning, Creation, Usage: Benchmarking LLMs for Comprehensive Tool Utilization in Real-World Complex Scenarios
ACL 2024
·
Shijue Huang
DBLP profile ↗
ORCID search ↗
ProxyQA: An Alternative Framework for Evaluating Long-Form Text Generation with Large Language Models
ACL 2024
·
Haochen Tan
DBLP profile ↗
ORCID search ↗
SINKT: A Structure-Aware Inductive Knowledge Tracing Model with Large Language Model
CIKM 2024
·
Lingyue Fu
DBLP profile ↗
ORCID search ↗