PPaperPicks

Hongning Wang

Tsinghua University, Department of Computer Science, Beijing, China

55 papers at tracked venues · 48 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Data Efficient RLVR via Off-Policy Influence Guidance
  2. Glyph: Scaling Context Windows via Visual-Text Compression
  3. HoWToBench: Holistic Evaluation for LLM's Capability in Human-level Writing using Tree of Writing
  4. How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study
  5. IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation
  6. IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation
  7. LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety
  8. When Smiley Turns Hostile: Interpreting How Emojis Trigger LLMs' Toxicity
  9. CharacterBench: Benchmarking Character Customization of Large Language Models
  10. CodePlan: Unlocking Reasoning Potential in Large Language Models by Scaling Code-form Planning
  11. Crisp: Cognitive Restructuring of Negative Thoughts through Multi-turn Supportive Dialogues
  12. Data Selection via Optimal Control for Language Models
  13. Decoupling Knowledge and Context: An Efficient and Effective Retrieval Augmented Generation Framework via Cross Attention
  14. Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints
  15. HPSS: Heuristic Prompting Strategy Search for LLM Evaluators
  16. JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
  17. Knowledge-to-Jailbreak: Investigating Knowledge-driven Jailbreaking Attacks for Large Language Models
  18. LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models
  19. LongSafety: Evaluating Long-Context Safety of Large Language Models
  20. MAPS: Advancing Multi-Modal Reasoning in Expert-Level Physical Science
  21. Parametric Retrieval Augmented Generation
  22. RecFlow: An Industrial Full Flow Recommendation Dataset
  23. SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
  24. SelfRACG: Enabling LLMs to Self-Express and Retrieve for Code Generation
  25. ShieldVLM: Safeguarding the Multimodal Implicit Toxicity via Deliberative Reasoning with LVLMs: ShieldVLM
  26. SocialEval: Evaluating Social Intelligence of Large Language Models
  27. SocialSim: Towards Socialized Simulation of Emotional Support Conversation
  28. Tree-KG: An Expandable Knowledge Graph Construction Framework for Knowledge-intensive Domains
  29. VPO: Aligning Text-to-Video Generation Models with Prompt Optimization
  30. AMOR: A Recipe for Building Adaptable Modular Knowledge Agents Through Process Feedback
  31. AlignBench: Benchmarking Chinese Alignment of Large Language Models
  32. AutoDetect: Towards a Unified Framework for Automated Weakness Detection in Large Language Models
  33. Benchmarking Complex Instruction-Following with Multiple Constraints Composition
  34. Black-Box Prompt Optimization: Aligning Large Language Models without Model Training
  35. CharacterGLM: Customizing Social Characters with Large Language Models
  36. CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation
  37. Defending Large Language Models Against Jailbreaking Attacks Through Goal Prioritization
  38. Federated Linear Contextual Bandits with Heterogeneous Clients
  39. Full Stage Learning to Rank: A Unified Framework for Multi-Stage Systems
  40. Human vs. Generative AI in Content Creation Competition: Symbiosis or Conflict?
  41. Incentivized Truthful Communication for Federated Bandits
  42. Language Model Decoding as Direct Metrics Optimization
  43. Learning Task Decomposition to Assist Humans in Competitive Programming
  44. Meta-Reinforcement Learning via Exploratory Task Clustering
  45. Mitigating Reward Overoptimization via Lightweight Uncertainty Estimation
  46. Proceedings of the 47th International ACM SIGIR Conference on Research and Development in Information Retrieval, SIGIR 2024, Washington DC, USA, July 14-18, 2024
  47. Retention Depolarization in Recommender System
  48. ShieldLM: Empowering LLMs as Aligned, Customizable and Explainable Safety Detectors
  49. Stealthy Adversarial Attacks on Stochastic Multi-Armed Bandits
  50. The Second Workshop on Large Language Models for Individuals, Groups, and Society
  51. Third Workshop on Personalization and Recommendations in Search (PaRiS)
  52. Towards Efficient Exact Optimization of Language Model Alignment
  53. Unveiling User Satisfaction and Creator Productivity Trade-Offs in Recommendation Platforms
  54. User Welfare Optimization in Recommender Systems with Competing Content Creators
  55. WSDM 2024 Workshop on Large Language Models for Individuals, Groups, and Society