PPaperPicks

Mengnan Du

37 papers at tracked venues · 21 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AdaJudge: Adaptive Multi-Perspective Judging for Reward Modeling
  2. AdaptiveK: Complexity-Driven Sparse Autoencoders for Interpretable Language Model Representations
  3. DeepSieve: Information Sieving via LLM-as-a-Knowledge-Router
  4. Denoising Concept Vectors with Sparse Autoencoders for Improved Language Model Steering
  5. FaithLM: Towards Faithful Explanations for Large Language Models
  6. FinCall-Surprise: A Large Scale Multi-modal Benchmark for Earning Surprise Prediction
  7. FinChart-Bench: Benchmarking Financial Chart Comprehension in Vision-Language Models
  8. Fine-Grained Interpretation of Political Opinions in Large Language Models
  9. KnowThyself: An Agentic Assistant for LLM Interpretability
  10. LLM Agents in Law: Taxonomy, Applications, and Challenges
  11. SAE-FiRE: Enhancing Earnings Surprise Predictions Through Sparse Autoencoder Feature Selection
  12. SAGE: An Agentic Explainer Framework for Interpreting SAE Features in Language Models
  13. A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models
  14. Beyond Input Activations: Identifying Influential Latents by Gradient Sparse Autoencoders
  15. Beyond Single Concept Vector: Modeling Concept Subspace in LLMs with Gaussian Distribution
  16. Comparative Analysis of Demonstration Selection Algorithms for In-Context Learning in Large Language Models (Student Abstract)
  17. Concept-Centric Token Interpretation for Vector-Quantized Generative Models
  18. Data-centric NLP Backdoor Defense from the Lens of Memorization
  19. Feature Extraction and Steering for Enhanced Chain-of-Thought Reasoning in Language Models
  20. From Commands to Prompts: LLM-based Semantic File System for AIOS
  21. Improving LLM Reasoning through Interpretable Role-Playing Steering
  22. Invisible Backdoor Attack against Self-supervised Learning
  23. Language Ranker: A Metric for Quantifying LLM Performance Across High and Low-Resource Languages
  24. Large Vision-Language Model Alignment and Misalignment: A Survey Through the Lens of Explainability
  25. Massive Values in Self-Attention Modules are the Key to Contextual Knowledge Understanding
  26. NPC-TAG: Node Prompts for Classification on Text Attributed Graphs with LLMs
  27. Physics-Informed Attention-Enhanced Fourier Neural Operator for Solar Magnetic Field Extrapolations
  28. SAE-SSV: Supervised Steering in Sparse Representation Spaces for Reliable Control of Language Models
  29. AI Driven Online Advertising: Market Design, Generative AI, and Ethics
  30. Data-Centric Explainable Debiasing for Improving Fairness in Pre-trained Language Models
  31. Enhancing Fairness in In-Context Learning: Prioritizing Minority Samples in Demonstrations
  32. Explaining Time Series via Contrastive and Locally Sparse Perturbations
  33. LawLLM: Law Large Language Model for the US Legal System
  34. Secure Your Model: An Effective Key Prompt Protection Mechanism for Large Language Models
  35. Strategic Demonstration Selection for Improved Fairness in LLM In-Context Learning
  36. TVE: Learning Meta-attribution for Transferable Vision Explainer
  37. The Impact of Reasoning Step Length on Large Language Models