PPaperPicks

Lichao Sun

Lehigh University, Bethlehem, PA, USA

56 papers at tracked venues · 39 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. 3D4D: An Interactive, Editable, 4D World Model via 3D Video Generation
  2. BLURR: A Boosted Low-Resource Inference for Vision-Language-Action Model
  3. Rethinking and Red-Teaming Protective Perturbation in Personalized Diffusion Models
  4. SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator
  5. Vision-MoR: Scaling Vision Transformer via Patch-Level Mixture-of-Recursions
  6. Analytic Energy-Guided Policy Optimization for Offline Reinforcement Learning
  7. BadToken: Token-level Backdoor Attacks to Multi-modal Large Language Models
  8. BadVLA: Towards Backdoor Attacks on Vision-Language-Action Models via Objective-Decoupled Optimization
  9. BenTo: Benchmark Reduction with In-Context Transferability
  10. Benchmarking Vision Language Model Unlearning via Fictitious Facial Identity Dataset
  11. Both Text and Images Leaked! A Systematic Analysis of Data Contamination in Multimodal LLM
  12. Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
  13. DataGen: Unified Synthetic Dataset Generation via Large Language Models
  14. FinLLM-B: When Large Language Models Meet Financial Breakout Trading
  15. First SFT, Second RL, Third UPT: Continual Improving Multi-Modal LLM Reasoning via Unsupervised Post-Training
  16. GUI-World: A Video Benchmark and Dataset for Multimodal GUI-oriented Understanding
  17. Jailbreaking LLMs Through Alignment Vulnerabilities in Out-of-Distribution Settings
  18. LlaVA-CoT: Let Vision Language Models Reason Step-By-Step
  19. LongLLaVA: Scaling Multi-modal LLMs to 1000 Images Efficiently via a Hybrid Architecture
  20. Merge Hijacking: Backdoor Attacks to Model Merging of Large Language Models
  21. MetaAgents: Large Language Model Based Agents for Decision-Making on Teaming
  22. SAMed-2: Selective Memory Enhanced Medical Segment Anything Model
  23. Tackling Continual Offline RL through Selective Weights Activation on Aligned Spaces
  24. Towards Building Model/Prompt-Transferable Attackers against Large Vision-Language Models
  25. TruthPrInt: Mitigating Large Vision-Language Models Object Hallucination via Latent Truthful-Guided Pre-Intervention
  26. XAttnMark: Learning Robust Audio Watermarking with Cross-Attention
  27. 1+1\textgreater2: Can Large Language Models Serve as Cross-Lingual Knowledge Aggregators?
  28. ACT-Diffusion: Efficient Adversarial Consistency Training for One-Step Diffusion Models
  29. Adapting Large Language Model for Cross-Subject Semantic Decoding from Video-Stimulated fMRI
  30. AlignBench: Benchmarking Chinese Alignment of Large Language Models
  31. CodeIP: A Grammar-Guided Multi-Bit Watermark for Large Language Models of Code
  32. Conditional Score-Based Diffusion Model for Cortical Thickness Trajectory Prediction
  33. Decision Mamba: Reinforcement Learning via Hybrid Selective Sequence Modeling
  34. Deep Efficient Private Neighbor Generation for Subgraph Federated Learning
  35. EditShield: Protecting Unauthorized Image Editing by Instruction-Guided Diffusion Models
  36. FedSecurity: A Benchmark for Attacks and Defenses in Federated Learning and Federated LLMs
  37. Frequency-Aware GAN for Imperceptible Transfer Attack on 3D Point Clouds
  38. From Creation to Clarification: ChatGPT's Journey Through the Fake News Quagmire
  39. GTBench: Uncovering the Strategic Reasoning Capabilities of LLMs via Game-Theoretic Evaluations
  40. HonestLLM: Toward an Honest and Helpful Large Language Model
  41. Improving Interpretation Faithfulness for Vision Transformers
  42. In-Context Decision Transformer: Reinforcement Learning via Hierarchical Chain-of-Thought
  43. LLM-as-a-Coauthor: Can Mixed Human-Written and Machine-Generated Text Be Detected?
  44. Lecture-style Tutorial: Towards Graph Foundation Models
  45. MLLM-as-a-Judge: Assessing Multimodal LLM-as-a-Judge with Vision-Language Benchmark
  46. Medical Image Synthesis via Fine-Grained Image-Text Alignment and Anatomy-Pathology Prompting
  47. MetaCloak: Preventing Unauthorized Subject-Driven Text-to-Image Diffusion-Based Synthesis via Meta-Learning
  48. MetaTool Benchmark for Large Language Models: Deciding Whether to Use Tools and Which to Use
  49. Pandora's Box: Towards Building Universal Attackers against Real-World Large Vision-Language Models
  50. Physical Backdoor: Towards Temperature-Based Backdoor Attacks in the Physical World
  51. Position: TrustLLM: Trustworthiness in Large Language Models
  52. ReTA: Recursively Thinking Ahead to Improve the Strategic Reasoning of Large Language Models
  53. Revisiting Gradient Pruning: A Dual Realization for Defending against Gradient Attacks
  54. SpecHub: Provable Acceleration to Multi-Draft Speculative Decoding
  55. Stable Unlearnable Example: Enhancing the Robustness of Unlearnable Examples via Stable Error-Minimizing Noise
  56. Virtual Context Enhancing Jailbreak Attacks with Special Token Injection