PPaperPicks

Ali Payani

30 papers at tracked venues · 21 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Benchmarking LLMs for Political Science: A United Nations Perspective
  2. FAMA: Failure-Aware Meta-Agentic Framework for Open-Source LLMs in Interactive Tool Use Environments
  3. Model Editing as a Double-Edged Sword: Steering Agent Behavior Toward Beneficence or Harm
  4. A Generic Framework for Conformal Fairness
  5. AutoSDT: Scaling Data-Driven Discovery Tasks Toward Open Co-Scientists
  6. Beyond Semantic Entropy: Boosting LLM Uncertainty Quantification with Pairwise Semantic Similarity
  7. Beyond Single Concept Vector: Modeling Concept Subspace in LLMs with Gaussian Distribution
  8. Can Knowledge Editing Really Correct Hallucinations?
  9. Deliberate Reasoning in Language Models as Structure-Aware Planning with an Accurate World Model
  10. Effective Training Data Synthesis for Improving MLLM Chart Understanding
  11. How Can Input Reformulation Improve Tool Usage Accuracy in a Complex Dynamic Environment? A Study on tau-bench
  12. InfantAgent-Next: A Multimodal Generalist Agent for Automated Computer Interaction
  13. Investigating the Shortcomings of LLMs in Step-by-Step Legal Reasoning
  14. Language Ranker: A Metric for Quantifying LLM Performance Across High and Low-Resource Languages
  15. Large Vision-Language Model Alignment and Misalignment: A Survey Through the Lens of Explainability
  16. MDBench: A Synthetic Multi-Document Reasoning Benchmark Generated with Knowledge Guidance
  17. SAE-SSV: Supervised Steering in Sparse Representation Spaces for Reliable Control of Language Models
  18. A Federated Stochastic Multi-level Compositional Minimax Algorithm for Deep AUC Maximization
  19. Can LLMs Reason in the Wild with Programs?
  20. Data-Efficient Contrastive Language-Image Pretraining: Prioritizing Data Quality over Quantity
  21. Few-shot Adaptation to Distribution Shifts By Mixing Source and Target Embeddings
  22. Harnessing the Power of Large Language Models for Natural Language to First-Order Logic Translation
  23. Large Language Models Can Learn Temporal Reasoning
  24. MACM: Utilizing a Multi-Agent System for Condition Mining in Solving Complex Mathematical Problems
  25. Neural Additive Tensor Decomposition for Sparse Tensors
  26. SketchQL Demonstration: Zero-shot Video Moment Querying with Sketches
  27. SketchQL: Video Moment Querying with a Visual Query Interface
  28. TEILP: Time Prediction over Knowledge Graphs via Logical Reasoning
  29. Temporal Inductive Logic Reasoning over Hypergraphs
  30. When is Tree Search Useful for LLM Planning? It Depends on the Discriminator