PPaperPicks

Haiyang Yu

Alibaba Group, China

16 papers at tracked venues · 12 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. EvoRoute: Experience-Driven Self-Routing LLM Agent Systems
  2. ExpSeek: Self-Triggered Experience Seeking for Web Agents
  3. MemPO: Self-Memory Policy Optimization for Long-Horizon Agents
  4. DEMO: Reframing Dialogue Interaction with Fine-grained Element Modeling
  5. DeepSolution: Boosting Complex Engineering Solution Design via Tree-based Exploration and Bi-point Thinking
  6. EIFBENCH: Extremely Complex Instruction Following Benchmark for Large Language Models
  7. IOPO: Empowering LLMs with Complex Instruction Following via Input-Output Preference Optimization
  8. On the Role of Attention Heads in Large Language Model Safety
  9. StructRAG: Boosting Knowledge Intensive Reasoning of LLMs via Inference-time Hybrid Information Structurization
  10. Transferable Post-training via Inverse Value Learning
  11. How Alignment and Jailbreak Work: Explain LLM Safety through Intermediate Hidden States
  12. Language Models are Super Mario: Absorbing Abilities from Homologous Models as a Free Lunch
  13. Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QA
  14. Preference Ranking Optimization for Human Alignment
  15. Self-Retrieval: End-to-End Information Retrieval with One Large Language Model
  16. SoFA: Shielded On-the-fly Alignment via Priority Rule Following