P
PaperPicks
Conferences
Ming Jin
Virginia Tech, Blacksburg, VA, USA
24 papers at tracked venues · 17 at CORE A* · active 2024–2025
DBLP profile ↗
ORCID 0000-0001-7909-4545 ↗
Venues
NeurIPS
×6
EMNLP
×5
ICML
×5
ICLR
×3
AAAI
×1
ACL
×1
CVPR
×1
ECML-PKDD
×1
NAACL
×1
Frequent coauthors
Bilgehan Sel
DBLP profile ↗
ORCID search ↗
×5
Myeongseob Ko
DBLP profile ↗
ORCID search ↗
×4
Shangding Gu
DBLP profile ↗
ORCID search ↗
×3
Hyunin Lee
DBLP profile ↗
ORCID search ↗
×2
Mohammad Beigi
DBLP profile ↗
ORCID search ↗
×2
Hoang Anh Just
DBLP profile ↗
ORCID search ↗
×1
Junyu Guo
DBLP profile ↗
ORCID search ↗
×1
Lanxiao Huang
DBLP profile ↗
ORCID search ↗
×1
Padmaksha Roy
DBLP profile ↗
ORCID search ↗
×1
Mahavir Dabas
DBLP profile ↗
ORCID search ↗
×1
Jianfeng He
DBLP profile ↗
ORCID search ↗
×1
Yi Zeng
DBLP profile ↗
ORCID search ↗
×1
Papers
A Black Swan Hypothesis: The Role of Human Irrationality in AI Safety
ICLR 2025
·
Hyunin Lee
DBLP profile ↗
ORCID search ↗
DiPT: Enhancing LLM Reasoning through Diversified Perspective-Taking
NAACL 2025
·
Hoang Anh Just
DBLP profile ↗
ORCID search ↗
Don't Trade Off Safety: Diffusion Regularization for Constrained Offline RL
NeurIPS 2025
·
Junyu Guo
DBLP profile ↗
ORCID search ↗
From Capabilities to Performance: Evaluating Key Functional Properties of LLM Architectures in Penetration Testing
EMNLP 2025
·
Lanxiao Huang
DBLP profile ↗
ORCID search ↗
Improving Novel Anomaly Detection with Domain-Invariant Latent Representations
ECML-PKDD 2025
·
Padmaksha Roy
DBLP profile ↗
ORCID search ↗
Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning
ICML 2025
·
Mahavir Dabas
DBLP profile ↗
ORCID search ↗
LLMs Can Plan Only If We Tell Them
ICLR 2025
·
Bilgehan Sel
DBLP profile ↗
ORCID search ↗
LLMs Can Reason Faster Only If We Let Them
ICML 2025
·
Bilgehan Sel
DBLP profile ↗
ORCID search ↗
Position: AI Safety Must Embrace an Antifragile Perspective
ICML 2025
·
Ming Jin
Probing Hidden Knowledge Holes in Unlearned LLMs
NeurIPS 2025
·
Myeongseob Ko
DBLP profile ↗
ORCID search ↗
Reinforcement Learning with Backtracking Feedback
NeurIPS 2025
·
Bilgehan Sel
DBLP profile ↗
ORCID search ↗
Retracing the Past: LLMs Emit Training Data When They Get Lost
EMNLP 2025
·
Myeongseob Ko
DBLP profile ↗
ORCID search ↗
Robust Gymnasium: A Unified Modular Benchmark for Robust Reinforcement Learning
ICLR 2025
·
Shangding Gu
DBLP profile ↗
ORCID search ↗
Sycophancy Mitigation Through Reinforcement Learning with Uncertainty-Aware Adaptive Reasoning Trajectories
EMNLP 2025
·
Mohammad Beigi
DBLP profile ↗
ORCID search ↗
Algorithm of Thoughts: Enhancing Exploration of Ideas in Large Language Models
ICML 2024
·
Bilgehan Sel
DBLP profile ↗
ORCID search ↗
Balance Reward and Safety Optimization for Safe Reinforcement Learning: A Perspective of Gradient Manipulation
AAAI 2024
·
Shangding Gu
DBLP profile ↗
ORCID search ↗
Boosting Alignment for Post-Unlearning Text-to-Image Generative Models
NeurIPS 2024
·
Myeongseob Ko
DBLP profile ↗
ORCID search ↗
Can We Trust the Performance Evaluation of Uncertainty Estimation Methods in Text Summarization?
EMNLP 2024
·
Jianfeng He
DBLP profile ↗
ORCID search ↗
Enhancing Efficiency of Safe Reinforcement Learning via Sample Manipulation
NeurIPS 2024
·
Shangding Gu
DBLP profile ↗
ORCID search ↗
Fairness-Aware Meta-Learning via Nash Bargaining
NeurIPS 2024
·
Yi Zeng
DBLP profile ↗
ORCID search ↗
InternalInspector I²: Robust Confidence Estimation in LLMs through Internal States
EMNLP 2024
·
Mohammad Beigi
DBLP profile ↗
ORCID search ↗
Pausing Policy Learning in Non-stationary Reinforcement Learning
ICML 2024
·
Hyunin Lee
DBLP profile ↗
ORCID search ↗
Skin-in-the-Game: Decision Making via Multi-Stakeholder Alignment in LLMs
ACL 2024
·
Bilgehan Sel
DBLP profile ↗
ORCID search ↗
The Mirrored Influence Hypothesis: Efficient Data Influence Estimation by Harnessing Forward Passes
CVPR 2024
·
Myeongseob Ko
DBLP profile ↗
ORCID search ↗