PPaperPicks

Anningzhe Gao

15 papers at tracked venues · 6 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Mamba Hawkes Process for Event Sequence Modeling
  2. APLOT: Robust Reward Modeling via Adaptive Preference Learning with Optimal Transport
  3. Add-One-In: Incremental Sample Selection for Large Language Models via a Choice-Based Greedy Paradigm
  4. Aligning Language Models Using Follow-up Likelihood as Reward Signal
  5. Atoxia: Red-teaming Large Language Models with Target Toxic Answers
  6. CoD, Towards an Interpretable Medical Agent using Chain of Diagnosis
  7. DRBO: Mitigating Short Board Effect via Dynamic Reward Balancing in Multi-reward LLM Optimization
  8. Huatuo-26M, a Large-scale Chinese Medical QA Dataset
  9. Intermediate Domain Alignment and Morphology Analogy for Patent-Product Image Retrieval
  10. LLMs for Mathematical Modeling: Towards Bridging the Gap between Natural and Mathematical Languages
  11. MLLM-Bench: Evaluating Multimodal LLMs with Per-sample Criteria
  12. Self-Instructed Derived Prompt Generation Meets In-Context Learning: Unlocking New Potential of Black-Box LLMs
  13. Unlocking LLMs' Self-Improvement Capacity with Autonomous Learning for Domain Adaptation
  14. OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
  15. Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale