PPaperPicks

Wei Ye

Peking University, National Engineering Research Center for Software Engineering, Beijing, China

46 papers at tracked venues · 38 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. ASKD: Reinforcement Learning-Style Knowledge Distillation with Quality-Adaptive Skewness
  2. Beyond Semantic Relevance: Counterfactual Risk Minimization for Robust Retrieval-Augmented Generation
  3. Chain of Evidence: Pixel-Level Visual Attribution for Iterative Retrieval-Augmented Generation
  4. Data Selection for Multi-turn Dialogue Instruction Tuning
  5. Instruction Data Selection via Answer Divergence
  6. Joint Optimization of Training Data and Policy in RLHF
  7. LeLoRA: Learnable Low-Rank Adaptation of Large Language Models
  8. Learning from Contrasts: Synthesizing Reasoning Paths from Diverse Search Trajectories
  9. Modeling Uncertainty Trends for Timely Retrieval in Dynamic RAG
  10. NeuroSym-Cal: Bridging the Reasoning-Execution Gap in Code Generation via Hierarchical Calibration
  11. Rethinking the Sampling Criteria in Reinforcement Learning for LLM Reasoning: A Competence-Difficulty Alignment Perspective
  12. Retrieval as Generation: A Unified Framework with Self-Triggered Information Planning
  13. Thinking Twice Makes Large Language Models Safer and More Helpful
  14. All-Optical Nonlinear Diffractive Deep Network for Ultrafast Image Denoising
  15. Boosting Resilience of Large Language Models through Causality-Driven Robust Optimization
  16. Can You Really Trust Code Copilot? Evaluating Large Language Models from a Code Security Perspective
  17. GETMusic: Generating Music Tracks with a Unified Representation and Diffusion Framework
  18. HaDeMiF: Hallucination Detection and Mitigation in Large Language Models
  19. MPL: Multiple Programming Languages with Large Language Models for Information Extraction
  20. Queries Are Not Alone: Clustering Text Embeddings for Video Search
  21. Reasoning Through Execution: Unifying Process and Outcome Rewards for Code Generation
  22. Robustness to Spurious Correlations via Dynamic Knowledge Transfer
  23. SAEMark: Steering Personalized Multilingual LLM Watermarks with Sparse Autoencoders
  24. SampleMix: A Sample-wise Pre-training Data Mixing Strategy by Coordinating Data Quality and Diversity
  25. Supportiveness-based Knowledge Rewriting for Retrieval-augmented Language Modeling
  26. SymDPO: Boosting In-Context Learning of Large Multimodal Models with Symbol Demonstration Direct Preference Optimization
  27. VLM-R³: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
  28. AutoSurvey: Large Language Models Can Automatically Write Surveys
  29. Boosting Model Resilience via Implicit Adversarial Data Augmentation
  30. Conv-Adapter: Exploring Parameter Efficient Transfer Learning for ConvNets
  31. Enhancing In-Context Learning via Implicit Demonstration Augmentation
  32. FreeEval: A Modular Framework for Trustworthy and Efficient Evaluation of Large Language Models
  33. Hal-Eval: A Universal and Fine-grained Hallucination Evaluation Framework for Large Vision Language Models
  34. Hallucination Augmented Contrastive Learning for Multimodal Large Language Model
  35. KIEval: A Knowledge-grounded Interactive Evaluation Framework for Large Language Models
  36. Labels Need Prompts Too: Mask Matching for Natural Language Understanding Tasks
  37. MaVEn: An Effective Multi-granularity Hybrid Visual Encoding Framework for Multimodal Large Language Model
  38. NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
  39. PURE: Aligning LLM via Pluggable Query Reformulation for Enhanced Helpfulness
  40. PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization
  41. RAGLAB: A Modular and Research-Oriented Unified Framework for Retrieval-Augmented Generation
  42. Refining Corpora from a Model Calibration Perspective for Chinese Spelling Correction
  43. SG-Bench: Evaluating LLM Safety Generalization Across Diverse Tasks and Prompt Types
  44. Supervised Knowledge Makes Large Language Models Better In-context Learners
  45. TiMix: Text-Aware Image Mixing for Effective Vision-Language Pre-training
  46. What Makes a Good Order of Examples in In-Context Learning