PPaperPicks

Tianwei Zhang

Nanyang Technological University, School of Computer Science and Engineering, Singapore

49 papers at tracked venues · 36 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. LeakDojo: Decoding the Leakage Threats of RAG Systems
  2. ShadeEdit: A Utility-Preserving and Defense-Evasive Knowledge Manipulation Attack in Federated LLMs
  3. A Benchmark for Semantic Sensitive Information in LLMs Outputs
  4. An Engorgio Prompt Makes Large Language Model Babble on
  5. An LLM-Empowered Adaptive Evolutionary Algorithm for Multi-Component Deep Learning Systems
  6. Automated Red Teaming for Text-to-Image Models Through Feedback-Guided Prompt Iteration with Vision-Language Models
  7. BSemiFL: Semi-supervised Federated Learning via a Bayesian Approach
  8. Cowpox: Towards the Immunity of VLM-based Multi-Agent Systems
  9. Detecting Perception-Based Attacks using Visual Odometry: Inconsistency Modeling and Checking on Robotic States
  10. Disco4D: Disentangled 4D Human Generation and Animation from a Single Image
  11. Exploring Multimodal Challenges in Toxic Chinese Detection: Taxonomy, Benchmark, and Findings
  12. GPT-NER: Named Entity Recognition via Large Language Models
  13. Mask Image Watermarking
  14. Masked Sensory-Temporal Attention for Sensor Generalization in Quadruped Locomotion
  15. Mind the Cost of Scaffold! Benign Clients May Even Become Accomplices of Backdoor Attack
  16. Model Supply Chain Poisoning: Backdooring Pre-trained Models via Embedding Indistinguishability
  17. Rethinking Key-Value Cache Compression Techniques for Large Language Model Serving
  18. Safe + Safe = Unsafe? Exploring How Safe Images Can Be Exploited to Jailbreak Large Vision-Language Models
  19. SceneTAP: Scene-Coherent Typographic Adversarial Planner against Vision-Language Models in Real-World Environments
  20. Speculating LLMs' Chinese Training Data Pollution from Their Tokens
  21. TRUST-VLM: Thorough Red-Teaming for Uncovering Safety Threats in Vision-Language Models
  22. Taught Well Learned Ill: Towards Distillation-conditional Backdoor Attack
  23. Towards Resilient Safety-driven Unlearning for Diffusion Models against Downstream Fine-tuning
  24. Transstratal Adversarial Attack: Compromising Multi-Layered Defenses in Text-to-Image Models
  25. Understanding the Dark Side of LLMs' Intrinsic Self-Correction
  26. Unified Locomotion Transformer with Simultaneous Sim-to-Real Transfer for Quadrupeds
  27. VideoShield: Regulating Diffusion-based Video Generation Models via Watermarking
  28. When Audio and Text Disagree: Revealing Text Bias in Large Audio-Language Models
  29. ART: Automatic Red-teaming for Text-to-Image Models to Protect Benign Users
  30. AquaLoRA: Toward White-box Protection for Customized Stable Diffusion Models via Watermark LoRA
  31. Backdoor Attacks with Input-Unique Triggers in NLP
  32. BadEdit: Backdooring Large Language Models by Model Editing
  33. Beware of Road Markings: A New Adversarial Patch Attack to Monocular Depth Estimation
  34. COSMIC: Compress Satellite Image Efficiently via Diffusion Compensation
  35. Course-Correction: Safety Alignment Using Synthetic Preferences
  36. EvilEdit: Backdooring Text-to-Image Diffusion Models in One Second
  37. FedCDA: Federated Learning with Cross-rounds Divergence-aware Aggregation
  38. FedDSE: Distribution-aware Sub-model Extraction for Federated Learning over Resource-constrained Devices
  39. FedNLR: Federated Learning with Neuron-wise Learning Rates
  40. Improving the Generalization of Unseen Crowd Behaviors for Reinforcement Learning based Local Motion Planners
  41. Model X-ray: Detecting Backdoored Models via Decision Boundary
  42. Off-dynamics Conditional Diffusion Planners
  43. Purifying Quantization-conditioned Backdoors via Layer-wise Activation Correction with Distribution Approximation
  44. Robust-Wide: Robust Watermarking Against Instruction-Driven Image Editing
  45. SAME: Sample Reconstruction against Model Extraction Attacks
  46. State Chrono Representation for Enhancing Generalization in Reinforcement Learning
  47. The Earth is Flat because...: Investigating LLMs' Belief towards Misinformation via Persuasive Conversation
  48. Walking in Others' Shoes: How Perspective-Taking Guides Large Language Models in Reducing Toxicity and Bias
  49. You Only Query Once: An Efficient Label-Only Membership Inference Attack