PPaperPicks

Kaidi Xu

29 papers at tracked venues · 20 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. COIN: Uncertainty-Guarding Selective Question Answering for Foundation Models with Provable Risk Guarantees
  2. Dialogue is Better Than Monologue: Instructing Meidcal LLMs via Strategic Conversations
  3. IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generation
  4. Safety Alignment of Large Language Models via Contrasting Safe and Harmful Distributions
  5. DiffZOO: A Purely Query-Based Black-Box Attack for Red-teaming Text-to-Image Generative Model via Zeroth Order Optimization
  6. DynaCode: A Dynamic Complexity-Aware Code Benchmark for Evaluating Large Language Models in Code Generation
  7. GuideLLM: Exploring LLM-Guided Conversation with Applications in Autobiography Interviewing
  8. MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models
  9. Not Just Text: Uncovering Vision Modality Typographic Threats in Image Generation Models
  10. Optimizing Robustness and Accuracy in Mixture of Experts: A Dual-Model Approach
  11. SConU: Selective Conformal Uncertainty in Large Language Models
  12. Sparse Neurons Carry Strong Signals of Question Ambiguity in LLMs
  13. Transfer Attack for Bad and Good: Explain and Boost Adversarial Transferability across Multimodal Large Language Models
  14. TruthPrInt: Mitigating Large Vision-Language Models Object Hallucination via Latent Truthful-Guided Pre-Intervention
  15. ACT-Diffusion: Efficient Adversarial Consistency Training for One-Step Diffusion Models
  16. An Efficient Membership Inference Attack for the Diffusion Model by Proximal Initialization
  17. Can Protective Perturbation Safeguard Personal Data from Being Exploited by Stable Diffusion?
  18. Caterpillar: A Pure-MLP Architecture with Shifted-Pillars-Concatenation
  19. ConU: Conformal Uncertainty in Large Language Models with Correctness Coverage Guarantees
  20. Decoding Compressed Trust: Scrutinizing the Trustworthiness of Efficient LLMs Under Compression
  21. Dynamic Adversarial Attacks on Autonomous Driving Systems
  22. E3: Ensemble of Expert Embedders for Adapting Synthetic Image Detectors to New Generators Using Limited Data
  23. GTBench: Uncovering the Strategic Reasoning Capabilities of LLMs via Game-Theoretic Evaluations
  24. NN4SysBench: Characterizing Neural Network Verification for Computer Systems
  25. Position: TrustLLM: Trustworthiness in Large Language Models
  26. ReTA: Recursively Thinking Ahead to Improve the Strategic Reasoning of Large Language Models
  27. Shifting Attention to Relevance: Towards the Predictive Uncertainty Quantification of Free-Form Large Language Models
  28. Stable Unlearnable Example: Enhancing the Robustness of Unlearnable Examples via Stable Error-Minimizing Noise
  29. Unveiling Typographic Deceptions: Insights of the Typographic Vulnerability in Large Vision-Language Models