PPaperPicks

Liang Lin

Sun Yat-sen University, Guangzhou, China

57 papers at tracked venues · 50 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Backdoor Collapse: Eliminating Unknown Threats Via Known Backdoor Aggregation In Language Models
    ACL 2026 · Liang Lin
  2. HearSay Benchmark: Do Audio LLMs Leak What They Hear?
  3. Hidden in the Noise: Unveiling Backdoors in Audio LLMs Alignment Through Latent Acoustic Pattern Triggers
    AAAI 2026 · Liang Lin
  4. Human-Centric Open-Future Task Discovery: Formulation, Benchmark, and Scalable Tree-Based Search
  5. Neural Chain-of-Thought Search: Searching the Optimal Reasoning Path to Enhance Large Language Models
  6. RSA-Bench: Benchmarking Audio Large Models in Real-World Acoustic Scenarios
  7. SEE: Signal Embedding Energy for Quantifying Noise Interference in Large Audio Language Models
  8. Similarity-aware Probabilistic Embeddings Modeling for Video-Text Retrieval
  9. Stable Language Guidance for Vision-Language-Action Models
  10. Visually-Guided Policy Optimization for Multimodal Reasoning
  11. 3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians
  12. AlphaAgent: LLM-Driven Alpha Mining with Regularized Exploration to Counteract Alpha Decay
  13. Are High-Quality AI-Generated Images More Difficult for Models to Detect?
  14. Beyond the Destination: A Novel Benchmark for Exploration-Aware Embodied Question Answering
  15. Can We Achieve Efficient Diffusion without Self-Attention? Distilling Self-Attention Into Convolutions
  16. Chain of Methodologies: Scaling Test Time Computation without Training
  17. Cool-Fusion: Fuse Large Language Models without Training
  18. Cross-modal Causal Relation Alignment for Video Question Grounding
  19. DAGSM: Disentangled Avatar Generation with GS-enhanced Mesh
  20. DART: Dual Adaptive Refinement Transfer for Open-Vocabulary Multi-Label Recognition
  21. DSPNet: Dual-vision Scene Perception for Robust 3D Question Answering
  22. Delving into Cascaded Instability: A Lipschitz Continuity View on Image Restoration and Object Detection Synergy
  23. DreamFuse: Adaptive Image Fusion with Diffusion Transformer
  24. Free-Moref: Instantly Multiplexing Context Perception Capabilities of Video-Mllms Within Single Inference
  25. HyperCRS: Hypergraph-Aware Multi-Grained Preference Learning to Burst Filter Bubbles in Conversational Recommendation System
  26. Language Models as Implicit Tree Search
  27. MiniLongBench: The Low-cost Long Context Understanding Benchmark for Large Language Models
  28. Monitoring Primitive Interactions During the Training of DNNs
  29. Quadratic Coreset Selection: Certifying and Reconciling Sequence and Token Mining for Efficient Instruction Tuning
  30. Reproducible Vision-Language Models Meet Concepts Out of Pre-Training
  31. Robust Egocentric Referring Video Object Segmentation via Dual-Modal Causal Intervention
  32. RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs
  33. SR-FoT: A Syllogistic-Reasoning Framework of Thought for Large Language Models Tackling Knowledge-based Reasoning Tasks
  34. Sim-DETR: Unlock DETR for Temporal Sentence Grounding
  35. Thinking Before You Speak: A Proactive Test-time Scaling Approach
  36. Towards Long-Horizon Vision-Language Navigation: Platform, Benchmark and Method
  37. Towards Understanding the Robustness of Diffusion-Based Purification: A Stochastic Perspective
  38. VTON 360: High-Fidelity Virtual Try-On from Any Viewing Direction
  39. Why Multi-Interest Fairness Matters: Hypergraph Contrastive Multi-Interest Learning for Fair Conversational Recommender System
  40. AlignMiF: Geometry-Aligned Multimodal Implicit Field for LiDAR-Camera Joint Synthesis
  41. AttNS: Attention-Inspired Numerical Solving For Limited Data Scenarios
  42. Decoder-Only LLMs are Better Controllers for Diffusion Models
  43. Diagnosing and Rectifying Fake OOD Invariance: A Restructured Causal Approach
  44. Diversity Matters: User-Centric Multi-Interest Learning for Conversational Movie Recommendation
  45. EVE: Efficient Vision-Language Pre-training with Masked Prediction and Modality-Aware MoE
  46. FacetCRS: Multi-Faceted Preference Learning for Pricking Filter Bubbles in Conversational Recommender System
  47. HyCoRec: Hypergraph-Enhanced Multi-Preference Learning for Alleviating Matthew Effect in Conversational Recommendation
  48. Kepler codebook
  49. Learning Adaptive Spatial Coherent Correlations for Speech-Preserving Facial Expression Manipulation
  50. Learning Background Prompts to Discover Implicit Knowledge for Open Vocabulary Object Detection
  51. Let's Think Outside the Box: Exploring Leap-of-Thought in Large Language Models with Creative Humor Generation
  52. MarvelOVD: Marrying Object Recognition and Vision-Language Models for Robust Open-Vocabulary Object Detection
  53. Mirror Gradient: Towards Robust Multimodal Recommender Systems via Exploring Flat Local Minima
  54. Mitigating Matthew Effect: Multi-Hypergraph Boosted Multi-Interest Self-Supervised Learning for Conversational Recommendation
  55. Self-Supervised Emotion Representation Disentanglement for Speech-Preserving Facial Expression Manipulation
  56. Stripe Observation Guided Inference Cost-Free Attention Mechanism
  57. WildVidFit: Video Virtual Try-On in the Wild via Image-Based Controlled Diffusion Models