PPaperPicks

Yuhang Guo

Beijing Institute of Technology, School of Computer Science and Technology, China

17 papers at tracked venues · 13 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Beyond Literal Mapping: Benchmarking and Improving Non-Literal Translation Evaluation
  2. Incorporating Self-Rewriting into Large Language Model Reasoning Reinforcement
  3. Mem²Evolve: Towards Self-Evolving Agents via Co-Evolutionary Capability Expansion and Experience Distillation
  4. PEAP: Proactive Embodied Action Sequence Planning with Joint Understanding of Vision and Audio Perception
  5. PEC-Home: Interpretation of Progressively Elliptical Commands in Smart Homes
  6. DocMEdit: Towards Document-Level Model Editing
  7. Exploring In-Image Machine Translation with Real-World Background
  8. HomeBench: Evaluating LLMs in Smart Homes with Valid and Invalid Instructions Across Single and Multiple Devices
  9. PRIM: Towards Practical In-Image Multilingual Machine Translation
  10. ReFF: Reinforcing Format Faithfulness in Language Models Across Varied Tasks
  11. RepoDebug: Repository-Level Multi-Task and Multi-Language Debugging Evaluation of Large Language Models
  12. SafeToolBench: Pioneering a Prospective Benchmark to Evaluating Tool Utilization Safety in LLMs
  13. ToolSpectrum: Towards Personalized Tool Utilization for Large Language Models
  14. TransBench: Breaking Barriers for Transferable Graphical User Interface Agents in Dynamic Digital Environments
  15. Deterministic Reversible Data Augmentation for Neural Machine Translation
  16. FAME: Towards Factual Multi-Task Model Editing
  17. Medical Dialogue System: A Survey of Categories, Methods, Evaluation and Challenges