P
PaperPicks
Conferences
Qiaozhi He
6 papers at tracked venues · 6 at CORE A* · active 2025–2026
DBLP profile ↗
ORCID search ↗
Venues
AAAI
×2
ACL
×2
CVPR
×1
ICML
×1
Frequent coauthors
Chenglong Wang
DBLP profile ↗
ORCID search ↗
×4
Yifu Huo
DBLP profile ↗
ORCID search ↗
×1
Tao Wang
DBLP profile ↗
ORCID search ↗
×1
Papers
Probing Preference Representations: A Multi-Dimensional Evaluation and Analysis Method for Reward Models
AAAI 2026
·
Chenglong Wang
DBLP profile ↗
ORCID search ↗
SERM: Self-Evolving Relevance Model with Agent-Driven Learning from Massive Query Streams
ACL 2026
·
Chenglong Wang
DBLP profile ↗
ORCID search ↗
SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models
ACL 2026
·
Yifu Huo
DBLP profile ↗
ORCID search ↗
GRAM: A Generative Foundation Reward Model for Reward Generalization
ICML 2025
·
Chenglong Wang
DBLP profile ↗
ORCID search ↗
RoVRM: A Robust Visual Reward Model Optimized via Auxiliary Textual Preference Data
AAAI 2025
·
Chenglong Wang
DBLP profile ↗
ORCID search ↗
StickMotion: Generating 3D Human Motions by Drawing a Stickman
CVPR 2025
·
Tao Wang
DBLP profile ↗
ORCID search ↗