PPaperPicks

Muyun Yang

Harbin Institute of Technology

17 papers at tracked venues · 12 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Diagnosing and Remedying Representation Deficiencies for Deterministic Reasoning in KGQA
  2. Long-form RewardBench: Evaluating Reward Models for Long-form Generation
  3. Lost in Benchmarks? Rethinking Large Language Model Benchmarking with Item Response Theory
  4. WSDPO: A Generative Word Sense Disambiguation Framework with Chain-of-Thought and Preference Optimization
  5. An Empirical Study of LLM-as-a-Judge for LLM Evaluation: Fine-tuned Judge Model is not a General Substitute for GPT-4
  6. Benchmarking LLMs for Translating Classical Chinese Poetry: Evaluating Adequacy, Fluency, and Elegance
  7. LLM-based Translation Inference with Iterative Bilingual Understanding
  8. Legal Fact Prediction: The Missing Piece in Legal Judgment Prediction
  9. Look Before You Leap: Enhance Attention and Vigilance Regarding Harmful Content with GuidelineLLM
  10. MADAWSD: Multi-Agent Debate Framework for Adversarial Word Sense Disambiguation
  11. Make Imagination Clearer! Stable Diffusion-based Visual Imagination for Multimodal Machine Translation
  12. Memory-augmented Query Reconstruction for LLM-based Knowledge Graph Reasoning
  13. MuSC: Improving Complex Instruction Following with Multi-granularity Self-Contrastive Training
  14. Thinking in Character: Advancing Role-Playing Agents with Role-Aware Reasoning
  15. DUAL-REFLECT: Enhancing Large Language Models for Reflective Translation through Dual Learning Feedback Mechanisms
  16. Dynamic Planning for LLM-based Graphical User Interface Automation
  17. Self-Evaluation of Large Language Model based on Glass-box Features