PPaperPicks

Soujanya Poria

Singapore University of Technology and Design, Singapore

35 papers at tracked venues · 19 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. 10 Open Challenges Steering the Future of Vision-Language-Action Models
    AAAI 2026 · Soujanya Poria
  2. DialogXpert: Driving Intelligent and Emotion-Aware Conversations Through Online Value-Based Reinforcement Learning with LLM Priors
  3. AlgoPuzzleVQA: Diagnosing Multimodal Reasoning Challenges of Language Models with Algorithmic Multimodal Puzzles
  4. DiffPO: Diffusion-styled Preference Optimization for Inference Time Alignment of Large Language Models
  5. Emma-X: An Embodied Multimodal Action Model with Grounded Chain of Thought and Look-ahead Spatial Reasoning
  6. Error Typing for Smarter Rewards: Improving Process Reward Models with Error-Aware Hierarchical Supervision
  7. Evaluating AI for Finance: Is AI Credible at Assessing Investment Risk Appetite?
  8. Evaluating LLMs' Mathematical and Coding Competency through Ontology-guided Interventions
  9. Ferret: Faster and Effective Automated Red Teaming with Reward-Based Scoring Technique
  10. From Grounding to Manipulation: Case Studies of Foundation Model Integration in Embodied Robotic Systems
  11. Libra-Leaderboard: Towards Responsible AI through a Balanced Leaderboard of Safety and Capability
  12. M-LongDoc: A Benchmark For Multimodal Super-Long Document Understanding And A Retrieval-Aware Tuning Framework
  13. MOOSE-Chem: Large Language Models for Rediscovering Unseen Chemistry Scientific Hypotheses
  14. Measuring and Enhancing Trustworthiness of LLMs in RAG through Grounded Attributions and Learning to Refuse
  15. PREMISE: Matching-based Prediction for Accurate Review Recommendation
  16. Pixel-Level Reasoning Segmentation via Multi-turn Conversations
  17. Reward-Guided Tree Search for Inference Time Alignment of Large Language Models
  18. The ACM Multimedia 2025 Grand Challenge of Multimodal Conversational Aspect-based Sentiment Analysis
  19. Why AI Is WEIRD and Shouldn't Be This Way: Towards AI for Everyone, with Everyone, by Everyone
  20. CM-TTS: Enhancing Real Time Text-to-Speech Synthesis Efficiency through Weighted Samplers and Consistency Models
  21. Chain-of-Knowledge: Grounding Large Language Models via Dynamic Knowledge Adapting over Heterogeneous Sources
  22. Consistency Guided Knowledge Retrieval and Denoising in LLMs for Zero-shot Document-level Relation Triplet Extraction
  23. Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
  24. Large Language Models for Automated Open-domain Scientific Hypotheses Discovery
  25. Mustango: Toward Controllable Text-to-Music Generation
  26. PanoSent: A Panoptic Sextuple Extraction Benchmark for Multimodal Conversational Aspect-based Sentiment Analysis
  27. PuzzleVQA: Diagnosing Multimodal Reasoning Challenges of Language Models with Abstract Visual Patterns
  28. Reasoning Paths Optimization: Learning to Reason and Explore From Diverse Paths
  29. Safety Arithmetic: A Framework for Test-time Safety Alignment of Language Models by Steering Parameters and Activations
  30. Self-Adaptive Sampling for Accurate Video Question Answering on Image Text Models
  31. Sowing the Wind, Reaping the Whirlwind: The Impact of Editing Language Models
  32. Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization
  33. Understanding the Capabilities and Limitations of Large Language Models for Cultural Commonsense
  34. Understanding, Leveraging, and Improving Large Language Models
    WWW 2024 · Soujanya Poria
  35. WalledEval: A Comprehensive Safety Evaluation Toolkit for Large Language Models