P
PaperPicks
Conferences
Yifu Huo
8 papers at tracked venues · 7 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
AAAI
×4
ACL
×2
EMNLP
×1
ICML
×1
Frequent coauthors
Chenglong Wang
DBLP profile ↗
ORCID search ↗
×6
Papers
GRAM-R²: Self-Training Generative Foundation Reward Models for Reward Reasoning
AAAI 2026
·
Chenglong Wang
DBLP profile ↗
ORCID search ↗
Probing Preference Representations: A Multi-Dimensional Evaluation and Analysis Method for Reward Models
AAAI 2026
·
Chenglong Wang
DBLP profile ↗
ORCID search ↗
SERM: Self-Evolving Relevance Model with Agent-Driven Learning from Massive Query Streams
ACL 2026
·
Chenglong Wang
DBLP profile ↗
ORCID search ↗
SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models
ACL 2026
·
Yifu Huo
GRAM: A Generative Foundation Reward Model for Reward Generalization
ICML 2025
·
Chenglong Wang
DBLP profile ↗
ORCID search ↗
HEAL: A Hypothesis-Based Preference-Aware Analysis Framework
EMNLP 2025
·
Yifu Huo
RoVRM: A Robust Visual Reward Model Optimized via Auxiliary Textual Preference Data
AAAI 2025
·
Chenglong Wang
DBLP profile ↗
ORCID search ↗
ESRL: Efficient Sampling-Based Reinforcement Learning for Sequence Generation
AAAI 2024
·
Chenglong Wang
DBLP profile ↗
ORCID search ↗