PPaperPicks

Kaixuan Huang

10 papers at tracked venues · 6 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. Deep Reinforcement Learning for Efficient and Fair Allocation of Healthcare Resources
  2. Emergent Symbolic Mechanisms Support Abstract Reasoning in Large Language Models
  3. MATH-Perturb: Benchmarking LLMs' Math Reasoning Abilities against Hard Perturbations
    ICML 2025 · Kaixuan Huang
  4. SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal
  5. Temporal Consistency for LLM Reasoning Process Error Identification
  6. Transformer-Based Multi-Agent Reinforcement Learning Method With Credit-Oriented Strategy Differentiation
    IROS 2025 · Kaixuan Huang
  7. TreeBoN: Enhancing Inference-Time Alignment with Speculative Tree-Search and Best-of-N Sampling
  8. A Theoretical Perspective for Speculative Decoding Algorithm
  9. Assessing the Brittleness of Safety Alignment via Pruning and Low-Rank Modifications
  10. Visual Adversarial Examples Jailbreak Aligned Large Language Models