PPaperPicks

Baolin Peng

22 papers at tracked venues · 16 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. LEDGER: Scaling Agentic Document Editing with Dependency-aware Graph Retrieval
  2. SynthAgent: Adapting Web Agents with Synthetic Supervision
  3. CollabLLM: From Passive Responders to Active Collaborators
  4. Decoder-Hybrid-Decoder Architecture for Efficient Reasoning with Long Generation
  5. ExACT: Teaching AI Agents to Explore with Reflective-MCTS and Exploratory Learning
  6. GUI-Actor: Coordinate-Free Visual Grounding for GUI Agents
  7. Iterative Nash Policy Optimization: Aligning LLMs with General Preferences via No-Regret Learning
  8. Latent Action Pretraining from Videos
  9. LiteSearch: Efficient Tree Search with Dynamic Exploration Budget for Math Reasoning
  10. Magma: A Foundation Model for Multimodal AI Agents
  11. Reinforcement Learning for Reasoning in Large Language Models with One Training Example
  12. Self-Tuning: Instructing LLMs to Effectively Acquire New Knowledge through Self-Teaching
  13. SimulatorArena: Are User Simulators Reliable Proxies for Multi-Turn Evaluation of AI Assistants?
  14. Evaluating the Instruction-Following Robustness of Large Language Models to Prompt Injection
  15. Improving LLM Generations via Fine-Grained Self-Endorsement
  16. Self-Alignment for Factuality: Mitigating Hallucinations in LLMs via Self-Evaluation
  17. Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models
  18. Self-Consistency Boosts Calibration for Math Reasoning
  19. Sub-Sentence Encoder: Contrastive Learning of Propositional Semantic Representations
  20. Teaching Language Models to Self-Improve through Interactive Demonstrations
  21. The Trickle-down Impact of Reward Inconsistency on RLHF
  22. Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing