PPaperPicks

Shouhong Ding

53 papers at tracked venues · 49 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. D²Pruner: Debiased Importance and Structural Diversity for MLLM Token Pruning
  2. GloTok: Global Perspective Tokenizer for Image Reconstruction and Generation
  3. LSAP-PV: High-Fidelity Palm Vein Image Synthesis via Layered Spectral Absorption Projection-Guided Diffusion Model
  4. TripleFDS: Triple Feature Disentanglement and Synthesis for Scene Text Editing
  5. A Generalizable Face Security Detection Method via Unified Texture and Semantic Feature Framework
  6. Aigi-Holmes: Towards Explainable and Generalizable AI-Generated Image Detection via Multimodal Large Language Models
  7. Antidote: A Unified Framework for Mitigating LVLM Hallucinations in Counterfactual Presupposition and Object Perception
  8. DITL2: Dual-Stage Invariance Transfer Learning for Generalizable Document Image Tampering Localization
  9. Data Synthesis with Diverse Styles for Face Recognition via 3DMM-Guided Diffusion
  10. Diff-Palm: Realistic Palmprint Generation with Polynomial Creases and Intra-Class Variation Controllable Diffusion Models
  11. Dual Data Alignment Makes AI-Generated Image Detector Easier Generalizable
  12. EyeSeg: An Uncertainty-Aware Eye Segmentation Framework for AR/VR
  13. From Enhancement to Understanding: Build a Generalized Bridge for Low-Light Vision via Semantically Consistent Unsupervised Fine-Tuning
  14. Fuse Before Transfer: Knowledge Fusion for Heterogeneous Distillation
  15. Generalizing Deepfake Video Detection with Plug-and-Play: Video-Level Blending and Spatiotemporal Adapter Tuning
  16. Guard Me If You Know Me: Protecting Specific Face-Identity from Deepfakes
  17. Instruct Where the Model Fails: Generative Data Augmentation via Guided Self-contrastive Fine-tuning
  18. Large Continual Instruction Assistant
  19. MS-DETR: Towards Effective Video Moment Retrieval and Highlight Detection by Joint Motion-Semantic Learning
  20. Orthogonal Subspace Decomposition for Generalizable AI-Generated Image Detection
  21. PVTree: Realistic and Controllable Palm Vein Generation for Recognition Tasks
  22. PiD: Generalized AI-Generated Images Detection with Pixelwise Decomposition Residuals
  23. ROD-MLLM: Towards More Reliable Object Detection in Multimodal Large Language Models
  24. SlerpFace: Face Template Protection via Spherical Linear Interpolation
  25. SonarGuard2: Ultrasonic Face Liveness Detection Based on Adaptive Doppler Effect Feature Extraction
  26. Stylized-Face: A Million-Level Stylized Face Dataset for Face Recognition
  27. Switchable Token-Specific Codebook Quantization For Face Image Compression
  28. ToVE: Efficient Vision-Language Learning via Knowledge Transfer from Vision Experts
  29. Towards Rationale-Answer Alignment of LVLMs via Self-Rationale Calibration
  30. UIFace: Unleashing Inherent Model Capabilities to Enhance Intra-Class Diversity in Synthetic Face Recognition
  31. Unified Adversarial Augmentation for Improving Palmprint Recognition
  32. VISA: Group-wise Visual Token Selection and Aggregation via Graph Summarization for Efficient MLLMs Inference
  33. AlignCLIP: Align Multi Domains of Texts Input for CLIP models with Object-IoU Loss
  34. Anchor-based Robust Finetuning of Vision-Language Models
  35. Bilateral Adaptive Cross-Modal Fusion Prompt Learning for CLIP
  36. DF40: Toward Next-Generation Deepfake Detection
  37. DiffusionFake: Enhancing Generalization in Deepfake Detection via Guided Stable Diffusion
  38. Domain-Hallucinated Updating for Multi-Domain Face Anti-spoofing
  39. Enhancing Tampered Text Detection Through Frequency Feature Fusion and Decomposition
  40. HDMixer: Hierarchical Dependency with Extendable Patch for Multivariate Time Series Forecasting
  41. ID3: Identity-Preserving-yet-Diversified Diffusion Models for Synthetic Face Recognition
  42. LaRE2: Latent Reconstruction Error Based Method for Diffusion-Generated Image Detection
  43. MmAP: Multi-Modal Alignment Prompt for Cross-Domain Multi-Task Learning
  44. Model Tailor: Mitigating Catastrophic Forgetting in Multi-modal Large Language Models
  45. PCE-Palm: Palm Crease Energy Based Two-Stage Realistic Pseudo-Palmprint Generation
  46. Privacy-Preserving Face Recognition Using Trainable Feature Subtraction
  47. Re-Thinking Data Availability Attacks Against Deep Neural Networks
  48. Rethinking Generalizable Face Anti-Spoofing via Hierarchical Prototype-Guided Distribution Refinement in Hyperbolic Space
  49. SAFE: Slow and Fast Parameter-Efficient Tuning for Continual Learning with Pre-Trained Models
  50. SDPose: Tokenized Pose Estimation via Circulation-Guide Self-Distillation
  51. Second Edition FRCSyn Challenge at CVPR 2024: Face Recognition Challenge in the Era of Synthetic Data
  52. TF-FAS: Twofold-Element Fine-Grained Semantic Guidance for Generalizable Face Anti-spoofing
  53. Test-Time Domain Generalization for Face Anti-Spoofing