P
PaperPicks
Conferences
Yuheng Zhang
13 papers at tracked venues · 13 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
NeurIPS
×4
ICLR
×2
SIGMOD
×2
AAAI
×1
ACL
×1
ICML
×1
VLDB
×1
WWW
×1
Frequent coauthors
Mengxiao Zhang
DBLP profile ↗
ORCID search ↗
×2
Yuxuan Chou
DBLP profile ↗
ORCID search ↗
×1
Mingyue Huo
DBLP profile ↗
ORCID search ↗
×1
Zhencan Peng
DBLP profile ↗
ORCID search ↗
×1
Chenlu Ye
DBLP profile ↗
ORCID search ↗
×1
Papers
Improving Deepfake Detection with Reinforcement Learning-Based Adaptive Data Augmentation
AAAI 2026
·
Yuxuan Chou
DBLP profile ↗
ORCID search ↗
LSHAlign: All-Pair Near-Duplicate Text Alignment via Locality-Sensitive Hashing
SIGMOD 2026
·
Yuheng Zhang
Near-Duplicate Text Alignment under Weighted Jaccard Similarity
VLDB 2026
·
Yuheng Zhang
TagSpeech: End-to-End Multi-Speaker ASR and Diarization with Fine-Grained Temporal Grounding
ACL 2026
·
Mingyue Huo
DBLP profile ↗
ORCID search ↗
Improving LLM General Preference Alignment via Optimistic Online Mirror Descent
NeurIPS 2025
·
Yuheng Zhang
Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
ICLR 2025
·
Yuheng Zhang
Noise Matters: Diffusion Model-based Urban Mobility Generation with Collaborative Noise Priors
WWW 2025
·
Yuheng Zhang
Statistical Tractability of Off-policy Evaluation of History-dependent Policies in POMDPs
ICLR 2025
·
Yuheng Zhang
Efficient Contextual Bandits with Uninformed Feedback Graphs
ICML 2024
·
Mengxiao Zhang
DBLP profile ↗
ORCID search ↗
Near-Duplicate Text Alignment with One Permutation Hashing
SIGMOD 2024
·
Zhencan Peng
DBLP profile ↗
ORCID search ↗
On the Curses of Future and History in Future-dependent Value Functions for Off-policy Evaluation
NeurIPS 2024
·
Yuheng Zhang
Online Iterative Reinforcement Learning from Human Feedback with General Preference Model
NeurIPS 2024
·
Chenlu Ye
DBLP profile ↗
ORCID search ↗
Provably Efficient Interactive-Grounded Learning with Personalized Reward
NeurIPS 2024
·
Mengxiao Zhang
DBLP profile ↗
ORCID search ↗