P
PaperPicks
Conferences
Dong Yu
Tencent AI Lab, China
84 papers at tracked venues · 50 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0003-0520-6844 ↗
Google Scholar ↗
Homepage ↗
Venues
ACL
×26
EMNLP
×14
InterSpeech
×12
NeurIPS
×9
ICLR
×6
NAACL
×6
AAAI
×5
ICML
×3
ACM MM
×1
EACL
×1
IJCAI
×1
Frequent coauthors
Ante Wang
DBLP profile ↗
ORCID search ↗
×4
Zhenwen Liang
DBLP profile ↗
ORCID search ↗
×3
Yebowen Hu
DBLP profile ↗
ORCID search ↗
×3
Mukai Li
DBLP profile ↗
ORCID search ↗
×2
Zhisong Zhang
DBLP profile ↗
ORCID search ↗
×2
Chenlong Deng
DBLP profile ↗
ORCID search ↗
×2
Yuheng Zhang
DBLP profile ↗
ORCID search ↗
×2
Ruihan Yang
DBLP profile ↗
ORCID search ↗
×2
Hongliang He
DBLP profile ↗
ORCID search ↗
×2
Manjie Xu
DBLP profile ↗
ORCID search ↗
×2
Yaoxun Xu
DBLP profile ↗
ORCID search ↗
×2
Ruixin Hong
DBLP profile ↗
ORCID search ↗
×2
Papers
Audio-Thinker: Guiding Large Audio Language Model When and How to Think via Reinforcement Learning
AAAI 2026
·
Shu Wu
DBLP profile ↗
ORCID search ↗
Crossing the Reward Bridge: Expanding Reinforcement Learning with Verifiable Rewards Across Diverse Domains
ACL 2026
·
Yi Su
DBLP profile ↗
ORCID search ↗
DegVoC: Revisiting Neural Vocoder from a Degradation Perspective
AAAI 2026
·
Andong Li
DBLP profile ↗
ORCID search ↗
EconProver: Towards More Economical Test-Time Scaling for Automated Theorem Proving
ACL 2026
·
Mukai Li
DBLP profile ↗
ORCID search ↗
Enhancing Stability and Fidelity for Zero-Shot TTS with a Multi-Level Evaluator
AAAI 2026
·
Hualei Wang
DBLP profile ↗
ORCID search ↗
Inference-Time Scaling of Verification: Self-Evolving Deep Research Agents via Test-Time Rubric-Guided Verification
ACL 2026
·
Yuxuan Wan
DBLP profile ↗
ORCID search ↗
Revisiting Audio-language Pretraining for Learning General-purpose Audio Representation
ACL 2026
·
Wei-Cheng Tseng
DBLP profile ↗
ORCID search ↗
Save the Good Prefix: Precise Error Penalization via Process-Supervised RL to Enhance LLM Reasoning
ACL 2026
·
Haolin Liu
DBLP profile ↗
ORCID search ↗
Too Correct to Learn: Reinforcement Learning on Saturated Reasoning Data
ACL 2026
·
Zhenwen Liang
DBLP profile ↗
ORCID search ↗
UniCUE: Unified Recognition and Generation Framework for Chinese Cued Speech Video-to-Speech Generation
AAAI 2026
·
Jinting Wang
DBLP profile ↗
ORCID search ↗
VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents
ACL 2026
·
Jiliang Hu
DBLP profile ↗
ORCID search ↗
Verified Critical Step Optimization for LLM Agents
ACL 2026
·
Mukai Li
DBLP profile ↗
ORCID search ↗
WebAggregator: Enhancing Compositional Reasoning Capabilities of Deep Research Agent Foundation Models
ACL 2026
·
Rui Wang
DBLP profile ↗
ORCID search ↗
WebRollback: Enhancing Web Agents with Explicit Rollback Mechanisms
EACL 2026
·
Zhisong Zhang
DBLP profile ↗
ORCID search ↗
Your Reasoning Model is Secretly a Reward Model - Optimization-Free Verification from Experience
ACL 2026
·
Zhenwen Liang
DBLP profile ↗
ORCID search ↗
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression
ACL 2025
·
Chenlong Deng
DBLP profile ↗
ORCID search ↗
Attention Entropy is a Key Factor: An Analysis of Parallel Context Encoding with Full-attention-based Pre-trained Language Models
ACL 2025
·
Zhisong Zhang
DBLP profile ↗
ORCID search ↗
BridgeVoC: Neural Vocoder with Schrödinger Bridge
IJCAI 2025
·
Tong Lei
DBLP profile ↗
ORCID search ↗
Cognitive Kernel: An Open-source Agent System towards Generalist Autopilots
NAACL 2025
·
Hongming Zhang
DBLP profile ↗
ORCID search ↗
DOTS: Learning to Reason Dynamically in LLMs via Optimal Reasoning Trajectories Search
ICLR 2025
·
Murong Yue
DBLP profile ↗
ORCID search ↗
DSBench: How Far Are Data Science Agents from Becoming Data Science Experts?
ICLR 2025
·
Liqiang Jing
DBLP profile ↗
ORCID search ↗
DeFine: Decision-Making with Analogical Reasoning over Factor Profiles
ACL 2025
·
Yebowen Hu
DBLP profile ↗
ORCID search ↗
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
EMNLP 2025
·
Zhaowei Wang
DBLP profile ↗
ORCID search ↗
Do NOT Think That Much for 2+3=? On the Overthinking of Long Reasoning Models
ICML 2025
·
Xingyu Chen
DBLP profile ↗
ORCID search ↗
Don't Get Lost in the Trees: Streamlining LLM Reasoning by Overcoming Tree Search Exploration Pitfalls
ACL 2025
·
Ante Wang
DBLP profile ↗
ORCID search ↗
Efficient Multilingual ASR Finetuning via LoRA Language Experts
InterSpeech 2025
·
Jiahong Li
DBLP profile ↗
ORCID search ↗
EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer
InterSpeech 2025
·
Jiarui Hai
DBLP profile ↗
ORCID search ↗
From Continuous to Discrete: Cross-Domain Collaborative General Speech Enhancement via Hierarchical Language Models
ACM MM 2025
·
Zhaoxi Mu
DBLP profile ↗
ORCID search ↗
Hearing from Silence: Reasoning Audio Descriptions from Silent Videos via Vision-Language Model
InterSpeech 2025
·
Yong Ren
DBLP profile ↗
ORCID search ↗
Improving LLM General Preference Alignment via Optimistic Online Mirror Descent
NeurIPS 2025
·
Yuheng Zhang
DBLP profile ↗
ORCID search ↗
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
ICLR 2025
·
Yuheng Zhang
DBLP profile ↗
ORCID search ↗
LeVo: High-Quality Song Generation with Multi-Preference Alignment
NeurIPS 2025
·
Shun Lei
DBLP profile ↗
ORCID search ↗
LiteSearch: Efficient Tree Search with Dynamic Exploration Budget for Math Reasoning
AAAI 2025
·
Ante Wang
DBLP profile ↗
ORCID search ↗
LoGU: Long-form Generation with Uncertainty Expressions
ACL 2025
·
Ruihan Yang
DBLP profile ↗
ORCID search ↗
LongMemEval: Benchmarking Chat Assistants on Long-Term Interactive Memory
ICLR 2025
·
Di Wu
DBLP profile ↗
ORCID search ↗
Low-Bit Quantization Favors Undertrained LLMs
ACL 2025
·
Xu Ouyang
DBLP profile ↗
ORCID search ↗
MPS-Prover: Advancing Stepwise Theorem Proving by Multi-Perspective Search and Data Curation
NeurIPS 2025
·
Zhenwen Liang
DBLP profile ↗
ORCID search ↗
Mitigating Audiovisual Mismatch in Visual-Guide Audio Captioning
InterSpeech 2025
·
Le Xu
DBLP profile ↗
ORCID search ↗
OpenWebVoyager: Building Multimodal Web Agents via Iterative Real-World Exploration, Feedback and Optimization
ACL 2025
·
Hongliang He
DBLP profile ↗
ORCID search ↗
Recall with Reasoning: Chain-of-Thought Distillation for Mamba's Long-Context Memory and Extrapolation
EMNLP 2025
·
Jun-Yu Ma
DBLP profile ↗
ORCID search ↗
RepoGraph: Enhancing AI Software Engineering with Repository-level Code Graph
ICLR 2025
·
Siru Ouyang
DBLP profile ↗
ORCID search ↗
Retrieval-augmented GUI Agents with Generative Guidelines
EMNLP 2025
·
Ran Xu
DBLP profile ↗
ORCID search ↗
Router-Tuning: A Simple and Effective Approach for Dynamic Depth
EMNLP 2025
·
Shwai He
DBLP profile ↗
ORCID search ↗
Scaling beyond Denoising: Submitted System and Findings in URGENT Challenge 2025
InterSpeech 2025
·
Zhihang Sun
DBLP profile ↗
ORCID search ↗
The First Few Tokens Are All You Need: An Efficient and Effective Unsupervised Prefix Fine-Tuning Method for Reasoning Models
NeurIPS 2025
·
Ke Ji
DBLP profile ↗
ORCID search ↗
Thoughts Are All Over the Place: On the Underthinking of Long Reasoning Models
NeurIPS 2025
·
Yue Wang
DBLP profile ↗
ORCID search ↗
Towards Diverse and Efficient Audio Captioning via Diffusion Models
InterSpeech 2025
·
Manjie Xu
DBLP profile ↗
ORCID search ↗
Trust, But Verify: A Self-Verification Approach to Reinforcement Learning with Verifiable Rewards
NeurIPS 2025
·
Xiaoyuan Liu
DBLP profile ↗
ORCID search ↗
Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional Training
NeurIPS 2025
·
Mengru Wang
DBLP profile ↗
ORCID search ↗
UNCLE: Benchmarking Uncertainty Expressions in Long-Form Generation
EMNLP 2025
·
Ruihan Yang
DBLP profile ↗
ORCID search ↗
UniGist: Towards General and Hardware-aligned Sequence-level Long Context Compression
NeurIPS 2025
·
Chenlong Deng
DBLP profile ↗
ORCID search ↗
Video-to-Audio Generation with Fine-grained Temporal Semantics
InterSpeech 2025
·
Yuchen Hu
DBLP profile ↗
ORCID search ↗
VietASR: Achieving Industry-level Vietnamese ASR with 50-hour labeled data and Large-Scale Speech Pretraining
InterSpeech 2025
·
Jianheng Zhuo
DBLP profile ↗
ORCID search ↗
WAKE: Watermarking Audio with Key Enrichment
InterSpeech 2025
·
Yaoxun Xu
DBLP profile ↗
ORCID search ↗
WebCoT: Enhancing Web Agent Reasoning by Reconstructing Chain-of-Thought in Reflection, Branching, and Rollback
EMNLP 2025
·
Minda Hu
DBLP profile ↗
ORCID search ↗
WebEvolver: Enhancing Web Agent Self-Improvement with Co-evolving World Model
EMNLP 2025
·
Tianqing Fang
DBLP profile ↗
ORCID search ↗
A Closer Look at the Self-Verification Abilities of Large Language Models in Logical Reasoning
NAACL 2024
·
Ruixin Hong
DBLP profile ↗
ORCID search ↗
Abstraction-of-Thought Makes Language Models Better Reasoners
EMNLP 2024
·
Ruixin Hong
DBLP profile ↗
ORCID search ↗
CLOMO: Counterfactual Logical Modification with Large Language Models
ACL 2024
·
Yinya Huang
DBLP profile ↗
ORCID search ↗
Chain-of-Note: Enhancing Robustness in Retrieval-Augmented Language Models
EMNLP 2024
·
Wenhao Yu
DBLP profile ↗
ORCID search ↗
Comparing Discrete and Continuous Space LLMs for Speech Recognition
InterSpeech 2024
·
Yaoxun Xu
DBLP profile ↗
ORCID search ↗
Dense X Retrieval: What Retrieval Granularity Should We Use?
EMNLP 2024
·
Tong Chen
DBLP profile ↗
ORCID search ↗
Fact-and-Reflection (FaR) Improves Confidence Calibration of Large Language Models
ACL 2024
·
Xinran Zhao
DBLP profile ↗
ORCID search ↗
From Language Modeling to Instruction Following: Understanding the Behavior Shift in LLMs after Instruction Tuning
NAACL 2024
·
Xuansheng Wu
DBLP profile ↗
ORCID search ↗
Generative Pre-trained Speech Language Model with Efficient Hierarchical Transformer
ACL 2024
·
Yongxin Zhu
DBLP profile ↗
ORCID search ↗
Improving LLM Generations via Fine-Grained Self-Endorsement
ACL 2024
·
Ante Wang
DBLP profile ↗
ORCID search ↗
InFoBench: Evaluating Instruction Following Ability in Large Language Models
ACL 2024
·
Yiwei Qin
DBLP profile ↗
ORCID search ↗
Learn Beyond The Answer: Training Language Models with Reflection for Mathematical Reasoning
EMNLP 2024
·
Zhihan Zhang
DBLP profile ↗
ORCID search ↗
MM-LLMs: Recent Advances in MultiModal Large Language Models
ACL 2024
·
Duzhen Zhang
DBLP profile ↗
ORCID search ↗
MMC: Advancing Multimodal Chart Understanding with Large-scale Instruction Tuning
NAACL 2024
·
Fuxiao Liu
DBLP profile ↗
ORCID search ↗
Make-A-Voice: Revisiting Voice Large Language Models as Scalable Multilingual and Multitask Learners
ACL 2024
·
Rongjie Huang
DBLP profile ↗
ORCID search ↗
Multi-Channel Multi-Speaker ASR Using Target Speaker's Solo Segment
InterSpeech 2024
·
Yiwen Shao
DBLP profile ↗
ORCID search ↗
Polarity Calibration for Opinion Summarization
NAACL 2024
·
Yuanyuan Lei
DBLP profile ↗
ORCID search ↗
Prompt-guided Precise Audio Editing with Diffusion Models
ICML 2024
·
Manjie Xu
DBLP profile ↗
ORCID search ↗
RIR-SF: Room Impulse Response Based Spatial Feature for Target Speech Recognition in Multi-Channel Multi-Speaker Scenarios
InterSpeech 2024
·
Yiwen Shao
DBLP profile ↗
ORCID search ↗
Rewards-in-Context: Multi-objective Alignment of Foundation Models with Dynamic Preference Adjustment
ICML 2024
·
Rui Yang
DBLP profile ↗
ORCID search ↗
Self-Consistency Boosts Calibration for Math Reasoning
EMNLP 2024
·
Ante Wang
DBLP profile ↗
ORCID search ↗
Skills-in-Context: Unlocking Compositionality in Large Language Models
EMNLP 2024
·
Jiaao Chen
DBLP profile ↗
ORCID search ↗
SportsMetrics: Blending Text and Numerical Data to Understand Information Fusion in LLMs
ACL 2024
·
Yebowen Hu
DBLP profile ↗
ORCID search ↗
Sub-Sentence Encoder: Contrastive Learning of Propositional Semantic Representations
NAACL 2024
·
Sihao Chen
DBLP profile ↗
ORCID search ↗
The Trickle-down Impact of Reward Inconsistency on RLHF
ICLR 2024
·
Lingfeng Shen
DBLP profile ↗
ORCID search ↗
Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing
NeurIPS 2024
·
Ye Tian
DBLP profile ↗
ORCID search ↗
WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models
ACL 2024
·
Hongliang He
DBLP profile ↗
ORCID search ↗
When Reasoning Meets Information Aggregation: A Case Study with Sports Narratives
EMNLP 2024
·
Yebowen Hu
DBLP profile ↗
ORCID search ↗