P
PaperPicks
Conferences
Lu Wang
Microsoft, Beijing, China
24 papers at tracked venues · 13 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0002-7305-1496 ↗
Google Scholar ↗
Homepage ↗
Venues
ACL
×9
EMNLP
×5
ICLR
×2
AAAI
×1
CIKM
×1
ECAI
×1
MLSys
×1
NAACL
×1
SIGKDD
×1
UAI
×1
WSDM
×1
Frequent coauthors
Xiusheng Huang
DBLP profile ↗
ORCID search ↗
×2
Kaikai An
DBLP profile ↗
ORCID search ↗
×2
Yue Chen
DBLP profile ↗
ORCID search ↗
×1
Qibin Wang
DBLP profile ↗
ORCID search ↗
×1
Zhiyuan Peng
DBLP profile ↗
ORCID search ↗
×1
Mengqi Liao
DBLP profile ↗
ORCID search ↗
×1
Junting Lu
DBLP profile ↗
ORCID search ↗
×1
Runchuan Zhu
DBLP profile ↗
ORCID search ↗
×1
Yudi Zhang
DBLP profile ↗
ORCID search ↗
×1
Chenghua Huang
DBLP profile ↗
ORCID search ↗
×1
Yichen Ouyang
DBLP profile ↗
ORCID search ↗
×1
Huawen Feng
DBLP profile ↗
ORCID search ↗
×1
Papers
Break Through the Compression Bottleneck: From Theory to Practice
ACL 2026
·
Xiusheng Huang
DBLP profile ↗
ORCID search ↗
DUET: Joint Exploration of User-Item Profiles in Recommendation System
ACL 2026
·
Yue Chen
DBLP profile ↗
ORCID search ↗
Learning to Refine: Self-Refinement of Parallel Reasoning in LLMs
ACL 2026
·
Qibin Wang
DBLP profile ↗
ORCID search ↗
RepoGenesis: Benchmarking End-to-End Microservice Generation from Readme to Repository
ACL 2026
·
Zhiyuan Peng
DBLP profile ↗
ORCID search ↗
Theory-optimal Quantization Based on Flatness
ACL 2026
·
Xiusheng Huang
DBLP profile ↗
ORCID search ↗
Zipage: Maintain High Request Concurrency for LLM Reasoning through Compressed PagedAttention
ACL 2026
·
Mengqi Liao
DBLP profile ↗
ORCID search ↗
AXIS: Efficient Human-Agent-Computer Interaction with API-First LLM-Based Agents
ACL 2025
·
Junting Lu
DBLP profile ↗
ORCID search ↗
AdaptFlow: Adaptive Workflow Optimization via Meta-Learning
EMNLP 2025
·
Runchuan Zhu
DBLP profile ↗
ORCID search ↗
ICL-Bandit: Relevance Labeling in Advertisement Recommendation Systems via LLM
EMNLP 2025
·
Lu Wang
LettinGo: Explore User Profile Generation for Recommendation System
SIGKDD 2025
·
Lu Wang
ProtoRAIL: A Risk-cognizant Imitation Agent for Adaptive vCPU Oversubscription In the Cloud
MLSys 2025
·
Lu Wang
RuAG: Learned-rule-augmented Generation for Large Language Models
ICLR 2025
·
Yudi Zhang
DBLP profile ↗
ORCID search ↗
Self-Evolved Reward Learning for LLMS
ICLR 2025
·
Chenghua Huang
DBLP profile ↗
ORCID search ↗
Thread: A Logic-Based Data Organization Paradigm for How-To Question Answering with Retrieval Augmented Generation
EMNLP 2025
·
Kaikai An
DBLP profile ↗
ORCID search ↗
Token-level Proximal Policy Optimization for Query Generation
EMNLP 2025
·
Yichen Ouyang
DBLP profile ↗
ORCID search ↗
WarriorCoder: Learning from Expert Battles to Augment Code Large Language Models
ACL 2025
·
Huawen Feng
DBLP profile ↗
ORCID search ↗
AutoRAG-HP: Automatic Online Hyper-Parameter Tuning for Retrieval-Augmented Generation
EMNLP 2024
·
Jia Fu
DBLP profile ↗
ORCID search ↗
Balance Reward and Safety Optimization for Safe Reinforcement Learning: A Perspective of Gradient Manipulation
AAAI 2024
·
Shangding Gu
DBLP profile ↗
ORCID search ↗
COIN: Chance-Constrained Imitation Learning for Safe and Adaptive Resource Oversubscription under Uncertainty
CIKM 2024
·
Lu Wang
Everything of Thoughts: Defying the Law of Penrose Triangle for Thought Generation
ACL 2024
·
Ruomeng Ding
DBLP profile ↗
ORCID search ↗
Interpretable Imitation Learning with Dynamic Causal Relations
WSDM 2024
·
Tianxiang Zhao
DBLP profile ↗
ORCID search ↗
Nissist: An Incident Mitigation Copilot based on Troubleshooting Guides
ECAI 2024
·
Kaikai An
DBLP profile ↗
ORCID search ↗
SELF-GUARD: Empower the LLM to Safeguard Itself
NAACL 2024
·
Zezhong Wang
DBLP profile ↗
ORCID search ↗
SMuCo: Reinforcement Learning for Visual Control via Sequential Multi-view Total Correlation
UAI 2024
·
Tong Cheng
DBLP profile ↗
ORCID search ↗