PPaperPicks

Soheil Feizi

University of Maryland (UMD), Reliable AI Lab, College Park, MD, USA

31 papers at tracked venues · 23 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Attacker's Noise Can Manipulate Your Audio-based LLM in the Real World
  2. Decomposition-Enhanced Training for Post-Hoc Attributions in Language Models
  3. Schoenfeld's Anatomy of Mathematical Reasoning by Language Models
  4. Your LLM Agents are Temporally Blind: The Misalignment Between Tool Use Decisions and Human Time Perception
  5. A Closer Look at Bias and Chain-of-Thought Faithfulness of Large (Vision) Language Models
  6. A Technical Report on "Erasing the Invisible": The 2024 NeurIPS Competition on Stress Testing Image Watermarks
  7. Adversarial Paraphrasing: A Universal Attack for Humanizing AI-Generated Text
  8. Almost AI, Almost Human: The Challenge of Detecting AI-Polished Writing
  9. DyePack: Provably Flagging Test Set Contamination in LLMs Using Backdoors
  10. How Learnable Grids Recover Fine Detail in Low Dimensions: A Neural Tangent Kernel Analysis of Multigrid Parametric Encodings
  11. Localizing Knowledge in Diffusion Transformers
  12. Mitigating Compositional Failures in Text-to-Image Models with Causal Text Embedding Refinement
  13. RePanda: Pandas-powered Tabular Verification and Reasoning
  14. Rethinking Artistic Copyright Infringements In the Era Of Text-to-Image Generative Models
  15. Tool Preferences in Agentic LLMs are Unreliable
  16. Understanding the Effect of using Semantically Meaningful Tokens for Visual Representation Learning
  17. Unearthing Skill-level Insights for Understanding Trade-offs of Foundation Models
  18. DRSM: De-Randomized Smoothing on Malware Classifier Providing Certified Robustness
  19. Decomposing and Interpreting Image Representations via Text in ViTs Beyond CLIP
  20. Distilling Knowledge from Text-to-Image Generative Models Improves Visio-Linguistic Reasoning in CLIP
  21. Fast Adversarial Attacks on Language Models In One GPU Minute
  22. IntCoOp: Interpretability-Aware Vision-Language Prompt Tuning
  23. LLM-Check: Investigating Detection of Hallucinations in Large Language Models
  24. Localizing and Editing Knowledge In Text-to-Image Generative Models
  25. Loki: Low-rank Keys for Efficient Sparse Attention
  26. Measuring Self-Supervised Representation Quality for Downstream Classification Using Discriminative Features
  27. On Mechanistic Knowledge Localization in Text-to-Image Generative Models
  28. PRIME: Prioritizing Interpretability in Failure Mode Extraction
  29. Robustness of AI-Image Detectors: Fundamental Limits and Practical Attacks
  30. Strong Baselines for Parameter-Efficient Few-Shot Fine-Tuning
  31. Understanding Information Storage and Transfer in Multi-Modal Large Language Models