PPaperPicks

Pin-Yu Chen

66 papers at tracked venues · 58 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Data-Driven Lipschitz Continuity: A Cost-Effective Approach to Improve Adversarial Robustness
  2. GRE Score: Generative Risk Evaluation for Large Language Models
  3. Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs
  4. Hey, That's My Data! Token-Only Dataset Inference in Large Language Models
  5. ImReasoner: Improving Memory-based Language Models for Reasoning-in-a-Haystack Tasks
  6. MegaCoin: Enhancing Medium-Grained Color Perception for Vision-Language Models
  7. OjaKV: Context-Aware Online Low-Rank KV Cache Compression
  8. RiskLab: A Controlled Toolkit for Probing Emergent Risks in LLM-Based Multi-Agent Systems
  9. Why LLM Safety Guardrails Collapse After Fine-tuning: A Similarity Analysis Between Alignment and Fine-tuning Datasets
  10. ZoomR: Memory Efficient Reasoning through Multi-Granularity Key Value Retrieval
  11. Adaptive Distraction: Probing LLM Contextual Robustness with Automated Tree Search
  12. Attention Tracker: Detecting Prompt Injection Attacks in LLMs
  13. CoP: Agentic Red-teaming for Large Language Models using Composition of Principles
  14. Combining Domain and Alignment Vectors Provides Better Knowledge-Safety Trade-offs in LLMs
  15. Defensive Prompt Patch: A Robust and Generalizable Defense of Large Language Models against Jailbreak Attacks
  16. Differentiable Prompt Learning for Vision Language Models
  17. DiffuseKronA: A Parameter Efficient Fine-tuning Method for Personalized Diffusion Models
  18. From PEFT to DEFT: Parameter Efficient Finetuning for Reducing Activation Density in Transformers
  19. Justice or Prejudice? Quantifying Biases in LLM-as-a-Judge
  20. Large Language Models can Become Strong Self-Detoxifiers
  21. PSBD: Prediction Shift Uncertainty Unlocks Backdoor Detection
  22. REFINE: Inversion-Free Backdoor Defense via Model Reprogramming
  23. Retention Score: Quantifying Jailbreak Risks for Vision Language Models
  24. Revisiting Mode Connectivity in Neural Networks with Bezier Surface
  25. SEAL: Safety-enhanced Aligned LLM Fine-tuning via Bilevel Data Selection
  26. SPARC: An AI-Based Speech Processing and Real-Time Correction System
  27. STAR: Spectral Truncation and Rescale for Model Merging
  28. Shape it Up! Restoring LLM Safety during Finetuning
  29. TabWak: A Watermark for Tabular Diffusion Models
  30. Token Highlighter: Inspecting and Mitigating Jailbreak Prompts for Large Language Models
  31. Training Nonlinear Transformers for Chain-of-Thought Inference: A Theoretical Generalization Analysis
  32. When is Task Vector Provably Effective for Model Editing? A Generalization Analysis of Nonlinear Transformers
  33. A Deep Dive into the Trade-Offs of Parameter-Efficient Preference Alignment Techniques
  34. A Provably Effective Method for Pruning Experts in Fine-tuned Sparse Mixture-of-Experts
    ICML 2024 ·
    Mohammed Nowaz Rabbani Chowdhury
  35. AutoVP: An Automated Visual Prompting Framework and Benchmark
  36. Be Your Own Neighborhood: Detecting Adversarial Examples by the Neighborhood Relations Built on Self-Supervised Learning
  37. Computational Complexity of Verifying the Group No-show Paradox
  38. Duwak: Dual Watermarks in Large Language Models
  39. Elijah: Eliminating Backdoors Injected in Diffusion Models via Distribution Shift
  40. Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!
  41. GREAT Score: Global Robustness Evaluation of Adversarial Perturbation using Generative Models
  42. Gradient Cuff: Detecting Jailbreak Attacks on Large Language Models by Exploring Refusal Loss Landscapes
  43. How Do Nonlinear Transformers Learn and Generalize in In-Context Learning?
  44. It's Never Too Late: Fusing Acoustic Information into Large Language Models for Automatic Speech Recognition
  45. Language Agnostic Code Embeddings
  46. Large Language Models are Efficient Learners of Noise-Robust Speech Recognition
  47. Larimar: Large Language Models with Episodic Memory Control
  48. Learning Optimal Projection for Forecast Reconciliation of Hierarchical Time Series
  49. Model Reprogramming: Resource-Efficient Cross-Domain Machine Learning
    AAAI 2024 · Pin-Yu Chen
  50. Navigating the Safety Landscape: Measuring Risks in Finetuning Large Language Models
  51. NeuralFuse: Learning to Recover the Accuracy of Access-Limited Neural Network Inference in Low-Voltage Regimes
  52. Overload: Latency Attacks on Object Detection for Edge Devices
  53. Position: TrustLLM: Trustworthiness in Large Language Models
  54. Prompting4Debugging: Red-Teaming Text-to-Image Diffusion Models by Finding Problematic Prompts
  55. Rethinking Backdoor Attacks on Dataset Distillation: A Kernel Method Perspective
  56. Revisiting Zeroth-Order Optimization for Memory-Efficient LLM Fine-Tuning: A Benchmark
  57. Ring-A-Bell! How Reliable are Concept Removal Methods For Diffusion Models?
  58. SF-DQN: Provable Knowledge Transfer using Successor Feature for Deep Reinforcement Learning
  59. Safe LoRA: The Silver Lining of Reducing Safety Risks when Finetuning Large Language Models
  60. Self-Taught Recognizer: Toward Unsupervised Adaptation for Speech Foundation Models
  61. SepsisLab: Early Sepsis Prediction with Uncertainty Quantification and Active Sensing
  62. The Devil is in the Neurons: Interpreting and Mitigating Social Biases in Language Models
  63. Time-LLM: Time Series Forecasting by Reprogramming Large Language Models
  64. Uncovering the Hidden Cost of Model Compression
  65. What Improves the Generalization of Graph Transformers? A Theoretical Dive into the Self-attention and Positional Encoding
  66. What Would Gauss Say About Representations? Probing Pretrained Image Models using Synthetic Gaussian Benchmarks