PPaperPicks

Shikun Zhang

41 papers at tracked venues · 33 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. ASKD: Reinforcement Learning-Style Knowledge Distillation with Quality-Adaptive Skewness
  2. Data Selection for Multi-turn Dialogue Instruction Tuning
  3. Instruction Data Selection via Answer Divergence
  4. Joint Optimization of Training Data and Policy in RLHF
  5. L2-LoRA: Improving Low-Rank Adaptation with Layer-Specific Regularization
  6. LeLoRA: Learnable Low-Rank Adaptation of Large Language Models
  7. Modeling Uncertainty Trends for Timely Retrieval in Dynamic RAG
  8. Rethinking the Sampling Criteria in Reinforcement Learning for LLM Reasoning: A Competence-Difficulty Alignment Perspective
  9. Retrieval as Generation: A Unified Framework with Self-Triggered Information Planning
  10. Thinking Twice Makes Large Language Models Safer and More Helpful
  11. ToolSafe: Enhancing Tool Invocation Safety of LLM-based agents via Proactive Step-level Guardrail and Feedback
  12. All-Optical Nonlinear Diffractive Deep Network for Ultrafast Image Denoising
  13. Boosting Resilience of Large Language Models through Causality-Driven Robust Optimization
  14. Can You Really Trust Code Copilot? Evaluating Large Language Models from a Code Security Perspective
  15. GETMusic: Generating Music Tracks with a Unified Representation and Diffusion Framework
  16. HaDeMiF: Hallucination Detection and Mitigation in Large Language Models
  17. MPL: Multiple Programming Languages with Large Language Models for Information Extraction
  18. Reasoning Through Execution: Unifying Process and Outcome Rewards for Code Generation
  19. Robustness to Spurious Correlations via Dynamic Knowledge Transfer
  20. SAEMark: Steering Personalized Multilingual LLM Watermarks with Sparse Autoencoders
  21. SampleMix: A Sample-wise Pre-training Data Mixing Strategy by Coordinating Data Quality and Diversity
  22. Supportiveness-based Knowledge Rewriting for Retrieval-augmented Language Modeling
  23. SymDPO: Boosting In-Context Learning of Large Multimodal Models with Symbol Demonstration Direct Preference Optimization
  24. VLM-R³: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought
  25. AutoSurvey: Large Language Models Can Automatically Write Surveys
  26. Boosting Model Resilience via Implicit Adversarial Data Augmentation
  27. Enhancing In-Context Learning via Implicit Demonstration Augmentation
  28. FreeEval: A Modular Framework for Trustworthy and Efficient Evaluation of Large Language Models
  29. Hal-Eval: A Universal and Fine-grained Hallucination Evaluation Framework for Large Vision Language Models
  30. Hallucination Augmented Contrastive Learning for Multimodal Large Language Model
  31. KIEval: A Knowledge-grounded Interactive Evaluation Framework for Large Language Models
  32. Labels Need Prompts Too: Mask Matching for Natural Language Understanding Tasks
  33. MaVEn: An Effective Multi-granularity Hybrid Visual Encoding Framework for Multimodal Large Language Model
  34. NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models
  35. PURE: Aligning LLM via Pluggable Query Reformulation for Enhanced Helpfulness
  36. PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization
  37. RAGLAB: A Modular and Research-Oriented Unified Framework for Retrieval-Augmented Generation
  38. Refining Corpora from a Model Calibration Perspective for Chinese Spelling Correction
  39. SG-Bench: Evaluating LLM Safety Generalization Across Diverse Tasks and Prompt Types
  40. TiMix: Text-Aware Image Mixing for Effective Vision-Language Pre-training
  41. What Makes a Good Order of Examples in In-Context Learning