PPaperPicks

Liwei Jiang

20 papers at tracked venues · 17 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. AI Debate Aids Assessment of Controversial Claims
  2. AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
  3. Artificial Hivemind: The Open-Ended Homogeneity of Language Models (and Beyond)
    NeurIPS 2025 · Liwei Jiang
  4. Can Language Models Reason about Individualistic Human Values and Preferences?
    ACL 2025 · Liwei Jiang
  5. CulturalBench: A Robust, Diverse and Challenging Benchmark for Measuring LMs' Cultural Knowledge Through Human-AI Red-Teaming
  6. DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life
  7. Guardrails and Security for LLMs: Safe, Secure and Controllable Steering of LLM Applications
  8. Online Covariance Estimation in Nonsmooth Stochastic Approximation
    COLT 2025 · Liwei Jiang
  9. Position: Political Neutrality in AI Is Impossible - But Here Is How to Approximate It
  10. SafetyAnalyst: Interpretable, Transparent, and Steerable Safety Moderation for AI Behavior
  11. To Err Is AI: A Case Study Informing LLM Flaw Reporting Practices
  12. Impossible Distillation for Paraphrasing and Summarization: How to Make High-quality Lemonade out of Small, Low-quality Model
  13. JAMDEC: Unsupervised Authorship Obfuscation using Constrained Decoding over Small Language Models
  14. MigraineTracker: Examining Patient Experiences with Goal-Directed Self-Tracking for a Chronic Health Condition
  15. Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
  16. Position: A Roadmap to Pluralistic Alignment
  17. The Generative AI Paradox: "What It Can Create, It May Not Understand"
  18. Value Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties
  19. WildGuard: Open One-stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
  20. WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models
    NeurIPS 2024 · Liwei Jiang