PPaperPicks

Jizhong Han

19 papers at tracked venues · 15 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Exploiting Synergistic Cognitive Biases to Bypass Safety in LLMs
  2. FABLE: Fine-grained Fact Anchoring for Unstructured Model Editing
  3. More Thinking, Less Talking: Internalizing Deliberative Safety into LLM Parameters
  4. Resolving the Security-Auditability Dilemma with Auditable Latent Chain-of-Thought Alignment
  5. Adaptive Test-Time Semantic Debiasing for AI-Generated Image Detection
  6. CatAID: Category-Guided AI-Generated Image Detection via Vision-Language Model Adaptation
  7. Chain of Attack: Hide Your Intention through Multi-Turn Interrogation
  8. Gamma-Guard: Lightweight Residual Adapters for Robust Guardrails in Large Language Models
  9. LyapLock: Bounded Knowledge Preservation in Sequential Large Language Model Editing
  10. OMS: One More Step Noise Searching to Enhance Membership Inference Attacks for Diffusion Models
  11. Resolution Attack: Exploiting Image Compression to Deceive Deep Neural Networks
  12. Unleashing the Temporal-Spatial Reasoning Capacity of GPT for Training-Free Audio and Language Referenced Video Object Segmentation
  13. Visual Perception Uncertainty Learning for Hallucination Detection in Large Vision-Language Models
  14. Customize your NeRF: Adaptive Source Driven 3D Scene Editing via Local-Global Iterative Training
  15. Dynamic Mixed-Prototype Model for Incremental Deepfake Detection
  16. Dynamic Prompting of Frozen Text-to-Image Diffusion Models for Panoptic Narrative Grounding
  17. Real Appearance Modeling for More General Deepfake Detection
  18. Uncertainty-Guided Modal Rebalance for Hateful Memes Detection
  19. Unveiling Structural Memorization: Structural Membership Inference Attack for Text-to-Image Diffusion Models