PPaperPicks

Nenghai Yu

University of Science and Technology of China, Hefei, China

47 papers at tracked venues · 38 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. EARG-Net: Edge-Aware Reconstruction-Guided Network for Image Manipulation Detection and Localization
  2. MagicPaint: Operate Anything for Image Inpainting with Diffusion Model
  3. When Agents Look the Same: Quantifying Distillation-Induced Similarity in Tool-Use Behaviors
  4. BinMetric: A Comprehensive Binary Code Analysis Benchmark for Large Language Models
  5. CompileAgent: Automated Real-World Repo-Level Compilation with Tool-Integrated LLM-based Agent System
  6. De-AntiFake: Rethinking the Protective Perturbations Against Voice Cloning Attacks
  7. Deciphering Cross-Modal Alignment in Large Vision-Language Models Via Modality Integration Rate
  8. EvoBench: Towards Real-world LLM-Generated Text Detection Benchmarking for Evolving Large Language Models
  9. FE-CLIP: Frequency Enhanced CLIP Model for Zero-Shot Anomaly Detection and Segmentation
  10. LD-RoViS: Training-free Robust Video Steganography for Deterministic Latent Diffusion Model
  11. MARS-Bench: A Multi-turn Athletic Real-world Scenario Benchmark for Dialogue Evaluation
  12. MES-RAG: Bringing Multi-modal, Entity-Storage, and Secure Enhancements to RAG
  13. MMPro: A Decoupled Perception-Thinking-Execution Framework for Secure GUI Agent
  14. Merging-Resistant Watermarking for LoRA Modules
  15. Mixture-of-Noises Enhanced Forgery-Aware Predictor for Multi-Face Manipulation Detection and Localization
  16. On the Vulnerability of Text Sanitization
  17. Rethinking Masked Data Reconstruction Pretraining for Strong 3D Action Representation Learning
  18. SQL Injection Jailbreak: A Structural Disaster of Large Language Models
  19. STEAD: Robust Provably Secure Linguistic Steganography with Diffusion Language Model
  20. Scale Your Instructions: Enhance the Instruction-Following Fidelity of Unified Image Generation Model by Self-Adaptive Attention Scaling
  21. StegoZip: Enhancing Linguistic Steganography Payload in Practice with Large Language Models
  22. T2SMark: Balancing Robustness and Diversity in Noise-as-Watermark for Diffusion Models
  23. TAG-WM: Tamper-Aware Generative Image Watermarking via Diffusion Inversion Sensitivity
  24. Towards Anytime Retrieval: A Benchmark for Anytime Person Re-Identification
  25. Towards Good Generalizations for Diffusion Generated Image Detection Using Multiple Reconstruction Contrastive Learning
  26. Training-free Open-Vocabulary Semantic Segmentation via Diverse Prototype Construction and Sub-region Matching
  27. UNICL-SAM: Uncertainty-Driven In-Context Segmentation with Part Prototype Discovery
  28. Vector Database Watermarking
  29. A Geometric Distortion Immunized Deep Watermarking Framework with Robustness Generalizability
  30. AquaLoRA: Toward White-box Protection for Customized Stable Diffusion Models via Watermark LoRA
  31. Boosting Vanilla Lightweight Vision Transformers via Re-parameterization
  32. DPIC: Decoupling Prompt and Intrinsic Characteristics for LLM Generated Text Detection
  33. Data-Free Hard-Label Robustness Stealing Attack
  34. FaceRSA: RSA-Aware Facial Identity Cryptography Framework
  35. Gaussian Shading: Provable Performance-Lossless Image Watermarking for Diffusion Models
  36. Llama SLayer 8B: Shallow Layers Hold the Key to Knowledge Injection
  37. Model X-ray: Detecting Backdoored Models via Decision Boundary
  38. MotionGPT: Finetuned LLMs Are General-Purpose Motion Generators
  39. MuST: Robust Image Watermarking for Multi-Source Tracing
  40. OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allocation
  41. ScalingFilter: Assessing Data Quality through Inverse Utilization of Scaling Laws
  42. SemGIR: Semantic-Guided Image Regeneration Based Method for AI-generated Image Detection and Attribution
  43. TCI-Former: Thermal Conduction-Inspired Transformer for Infrared Small Target Detection
  44. Text Fluoroscopy: Detecting LLM-Generated Text through Intrinsic Features
  45. Towards More Unified In-Context Visual Understanding
  46. Transferable Facial Privacy Protection against Blind Face Restoration via Domain-Consistent Adversarial Obfuscation
  47. Unifying Multi-Modal Uncertainty Modeling and Semantic Alignment for Text-to-Image Person Re-identification