PPaperPicks

Di Wang

King Abdullah University of Science and Technology (KAUST), Saudi Arabia

48 papers at tracked venues · 39 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AutoMonitor-Bench: Evaluating the Reliability of LLM-Based Misbehavior Monitor
  2. CoLA: A Choice Leakage Attack Framework to Expose Privacy Risks in Subset Training
  3. Curriculum-RLAIF: Curriculum Alignment with Reinforcement Learning from AI Feedback
  4. Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs
  5. High-Throughput and Memory-Efficient Zeroth-Order Fine-tuning LLMs with Distributed Parallel Computing
  6. PIXEL: Adaptive Steering Via Position-wise Injection with eXact Estimated Levels under a Subspace Calibration
  7. Understanding and Mitigating Political Stance Cross-topic Generalization in Large Language Models
  8. Visual Self-Fulfilling Alignment: Shaping Safety-Oriented Personas via Threat-Related Images
  9. When Truth Is Overridden: Uncovering the Internal Origins of Sycophancy in Large Language Models
  10. ABNet: Mitigating Sample Imbalance in Anomaly Detection Within Dynamic Graphs
  11. CODEMENV: Benchmarking Large Language Models on Code Migration
  12. COMPKE: Complex Question Answering under Knowledge Editing
  13. Can Large Language Models Identify Implicit Suicidal Ideation? An Empirical Evaluation
  14. Differentially Private Sparse Linear Regression with Heavy-Tailed Responses
  15. EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification
  16. Editable Concept Bottleneck Models
  17. Fair Text-to-Image Diffusion via Fair Mapping
  18. Fraud-R1 : A Multi-Round Benchmark for Assessing the Robustness of LLM Against Augmented Fraud and Phishing Inducements
  19. Improved Rates of Differentially Private Nonconvex-Strongly-Concave Minimax Optimization
  20. LUSTER: Link Prediction Utilizing Shared-Latent Space Representation in Multi-Layer Networks
  21. Locate-then-edit for Multi-hop Factual Recall under Knowledge Editing
  22. Mechanistic Unveiling of Transformer Circuits: Self-Influence as a Key to Model Reasoning
  23. Nearly Optimal Differentially Private ReLU Regression
  24. Privacy-Preserving Low-Rank Adaptation Against Membership Inference Attacks for Latent Diffusion Models
  25. Private Training Large-scale Models with Efficient DP-SGD
  26. Second-Order Convergence in Private Stochastic Non-Convex Optimization
  27. Semi-Supervised Concept Bottleneck Models
  28. Short-length Adversarial Training Helps LLMs Defend Long-length Jailbreak Attacks: Theoretical and Empirical Evidence
  29. Stable Vision Concept Transformers for Medical Diagnosis
  30. TraffiDent: A Dataset for Understanding the Interplay Between Traffic Dynamics and Incidents
  31. Understanding How Value Neurons Shape the Generation of Specified Values in LLMs
  32. Understanding the Repeat Curse in Large Language Models from a Feature Perspective
  33. An LLM can Fool Itself: A Prompt-Based Adversarial Attack
  34. Autonomous Workflow for Multimodal Fine-Grained Training Assistants Towards Mixed Reality
  35. Closing the Gap: Achieving Global Convergence (Last Iterate) of Actor-Critic under Markovian Sampling with Neural Network Parametrization
  36. Communication Efficient and Provable Federated Unlearning
  37. Dissecting Fine-Tuning Unlearning in Large Language Models
  38. Faithful Vision-Language Interpretation via Concept Bottleneck Models
  39. Improved Analysis of Sparse Linear Regression in Local Differential Privacy Model
  40. Improving Interpretation Faithfulness for Vision Transformers
  41. Perplexity-aware Correction for Robust Alignment with Noisy Preferences
  42. Privacy Amplification via Shuffling: Unified, Simplified, and Tightened
  43. Private Language Models via Truncated Laplacian Mechanism
  44. Revisiting Differentially Private ReLU Regression
  45. Theoretical Analysis of Robust Overfitting for Wide DNNs: An NTK Approach
  46. Towards Multi-dimensional Explanation Alignment for Medical Classification
  47. Truthful High Dimensional Sparse Linear Regression
  48. Understanding Forgetting in Continual Learning with Linear Regression