PPaperPicks

Anh Tuan Luu

68 papers at tracked venues · 45 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. From Stimuli to Minds: Enhancing Psychological Reasoning in LLMs via Bilateral Reinforcement Learning
  2. Learning Uncertainty from Sequential Internal Dispersion in Large Language Models
  3. MUR: Momentum Uncertainty guided Reasoning for Large Language Models
  4. P2P: A Poison-to-Poison Remedy for Reliable Backdoor Defense in LLMs
  5. Read as You See: Guiding Unimodal LLMs for Low-Resource Explainable Harmful Meme Detection
  6. Rethinking Reasoning: A Survey on Reasoning-based Backdoors in LLMs
  7. Rewarding the Rare: Uniqueness-Aware RL for Creative Problem Solving in LLMs
  8. Three Minds, One Legend: Jailbreak Large Reasoning Model with Adaptive Stacked Ciphers
  9. Towards Fast and Accurate Modeling for Cross-Lingual Label Projection
  10. Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization
  11. Afterburner: Reinforcement Learning Facilitates Self-Improving Code Efficiency Optimization
  12. AntiLeakBench: Preventing Data Contamination by Automatically Constructing Benchmarks with Updated Real-World Knowledge
  13. As Simple as Fine-tuning: LLM Alignment via Bidirectional Negative Feedback Loss
  14. Beyond In-Context Learning: Aligning Long-form Generation of Large Language Models via Task-Inherent Attribute Guidelines
  15. ClozeMath: Improving Mathematical Reasoning in Language Models by Learning to Fill Equations
  16. CodeArena: A Collective Evaluation Platform for LLM Code Generation
  17. Diffusion vs. Autoregressive Language Models: A Text Embedding Perspective
  18. Discrete Diffusion Language Model for Efficient Text Summarization
  19. EffiBench-X: A Multi-Language Benchmark for Measuring Efficiency of LLM-Generated Code
  20. Enhancing Multimodal Entity Linking with Jaccard Distance-based Conditional Contrastive Learning and Contextual Visual Augmentation
  21. FineReason: Evaluating and Improving LLMs' Deliberate Reasoning through Reflective Puzzle Solving
  22. Full-Step-DPO: Self-Supervised Preference Optimization with Step-wise Rewards for Mathematical Reasoning
  23. GeoPQA: Bridging the Visual Perception Gap in MLLMs for Geometric Reasoning
  24. HyperGraphRAG: Retrieval-Augmented Generation via Hypergraph-Structured Knowledge Representation
  25. Is Translation All You Need? A Study on Solving Multilingual Tasks with Large Language Models
  26. KBQA-o1: Agentic Knowledge Base Question Answering with Monte Carlo Tree Search
  27. LongRecipe: Recipe for Efficient Long Context Generalization in Large Language Models
  28. MRAG: A Modular Retrieval Framework for Time-Sensitive Question Answering
  29. Massively Multilingual Instruction-Following Information Extraction
  30. Motion-aware Contrastive Learning for Temporal Panoptic Scene Graph Generation
  31. Multi-Scale Contrastive Learning for Video Temporal Grounding
  32. SCOPE: Compress Mathematical Reasoning Steps for Efficient Automated Process Annotation
  33. SeaExam and SeaBench: Benchmarking LLMs with Local Multilingual Questions in Southeast Asia
  34. Unlearning Backdoor Attacks for LLMs with Weak-to-Strong Knowledge Distillation
  35. Unsupervised Hallucination Detection by Inspecting Reasoning Processes
  36. AKEW: Assessing Knowledge Editing in the Wild
  37. Are LLMs Good Zero-Shot Fallacy Classifiers?
  38. ChatKBQA: A Generate-then-Retrieve Framework for Knowledge Base Question Answering with Fine-tuned Large Language Models
  39. Data Augmentation using LLMs: Data Perspectives, Learning Paradigms and Challenges
  40. Defending Against Weight-Poisoning Backdoor Attacks for Parameter-Efficient Fine-Tuning
  41. Don't Forget Your Reward Values: Language Model Alignment via Value-based Calibration
  42. Encoding and Controlling Global Semantics for Long-form Video Question Answering
  43. Exploring the Potential of Large Language Models in Computational Argumentation
  44. Extractive Summarization with Text Generator
  45. FASTopic: Pretrained Transformer is a Fast, Adaptive, Stable, and Transferable Topic Model
  46. From Static to Dynamic: Knowledge Metabolism for Large Language Models
  47. Historical Embedding-Guided Efficient Large-Scale Federated Graph Learning
  48. KDMCSE: Knowledge Distillation Multimodal Sentence Embeddings with Adaptive Angular margin Contrastive Learning
  49. LAMPAT: Low-Rank Adaption for Multilingual Paraphrasing Using Adversarial Training
  50. Mercury: A Code Efficiency Benchmark for Code Large Language Models
  51. Meta-optimized Angular Margin Contrastive Framework for Video-Language Representation Learning
  52. Modeling Dynamic Topics in Chain-Free Fashion by Evolution-Tracking Contrastive Learning and Unassociated Word Exclusion
  53. Multi-expert Prompting Improves Reliability, Safety and Usefulness of Large Language Models
  54. On the Affinity, Rationality, and Diversity of Hierarchical Topic Modeling
  55. PoetryDiffusion: Towards Joint Semantic and Metrical Manipulation in Poetry Generation
  56. READ-PVLA: Recurrent Adapter with Partial Video-Language Alignment for Parameter-Efficient Transfer Learning in Low-Resource Video-Language Modeling
  57. Reasoning Paths Optimization: Learning to Reason and Explore From Diverse Paths
  58. SemRoDe: Macro Adversarial Training to Learn Representations that are Robust to Word-Level Attacks
  59. SynTQA: Synergistic Table-based Question Answering via Mixture of Text-to-SQL and E2E TQA
  60. Text2NKG: Fine-Grained N-ary Relation Extraction for N-ary relational Knowledge Graph Construction
  61. ToXCL: A Unified Framework for Toxic Speech Detection and Explanation
  62. Topic Modeling as Multi-Objective Contrastive Optimization
  63. Towards the TopMost: A Topic Modeling System Toolkit
  64. Uncertainty of Thoughts: Uncertainty-Aware Planning Enhances Information Seeking in LLMs
  65. UniBridge: A Unified Approach to Cross-Lingual Transfer Learning for Low-Resource Languages
  66. Universal Vulnerabilities in Large Language Models: Backdoor Attacks for In-context Learning
  67. Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives
  68. Who's Who: Large Language Models Meet Knowledge Conflicts in Practice