PPaperPicks

Jiangning Zhang

51 papers at tracked venues · 44 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Disco-RAG: Discourse-Aware Retrieval-Augmented Generation
  2. LLM-Oriented Token-Adaptive Knowledge Distillation
  3. UltraGen: High-Resolution Video Generation with Hierarchical Attention
  4. A Comprehensive Library for Benchmarking Multi-Class Visual Anomaly Detection
    ICCV 2025 · Jiangning Zhang
  5. AdaVideoRAG: Omni-Contextual Adaptive Retrieval-Augmented Efficient Long Video Understanding
  6. Are They the Same? Exploring Visual Correspondence Shortcomings of Multimodal LLMs
  7. Bridge Feature Matching and Cross-Modal Alignment with Mutual-Filtering for Zero-Shot Anomaly Detection
  8. CustAny: Customizing Anything from A Single Example
  9. Decouple and Track: Benchmarking and Improving Video Diffusion Transformers for Motion Transfer
  10. Explore In-Context Segmentation via Latent Diffusion Models
  11. GroundingFace: Fine-grained Face Understanding via Pixel Grounding Multimodal Large Language Model
  12. ID-Sculpt: ID-aware 3D Head Generation from Single In-the-wild Portrait Image
  13. Identity-Preserving Text-to-Video Generation Guided by Simple yet Effective Spatial-Temporal Decoupled Representations
  14. Improving Autoregressive Visual Generation with Cluster-Oriented Token Prediction
  15. LLaVA-KD: A Framework of Distilling Multimodal Large Language Models
  16. MobileMamba: Lightweight Multi-Receptive Visual Mamba Network
  17. OSV: One Step is Enough for High-Quality Image to Video Generation
  18. Point Cloud Mamba: Point Cloud Learning via State Space Model
  19. PointRWKV: Efficient RWKV-Like Model for Hierarchical Point Cloud Learning
  20. PointSeg: A Training-Free Paradigm for 3D Scene Segmentation via Foundation Models
  21. PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement
  22. Real-IAD D3: A Real-World 2D/Pseudo-3D/3D Dataset for Industrial Anomaly Detection
  23. SVFR: A Unified Framework for Generalized Video Face Restoration
  24. SaRA: High-Efficient Diffusion Model Fine-tuning with Progressive Sparse Low-Rank Adaptation
  25. Sonic: Shifting Focus to Global Audio Perception in Portrait Animation
  26. StrandDesigner: Towards Practical Strand Generation with Sketch Guidance
  27. TIMotion: Temporal and Interactive Framework for Efficient Human-Human Motion Generation
  28. UltraVideo: High-Quality UHD Video Dataset with Comprehensive Captions
  29. Unveil Inversion and Invariance in Flow Transformer for Versatile Image Editing
  30. A Diffusion-Based Framework for Multi-Class Anomaly Detection
  31. AdaCLIP: Adapting CLIP with Hybrid Learnable Prompts for Zero-Shot Anomaly Detection
  32. AnomalyDiffusion: Few-Shot Anomaly Image Generation with Diffusion Model
  33. COMD: Training-free Video Motion Transfer With Camera-Object Motion Disentanglement
  34. DiffuMatting: Synthesizing Arbitrary Objects with Matting-Level Annotation
  35. Face-Adapter for Pre-trained Diffusion Models with Fine-Grained ID and Attribute Control
  36. Fetch and Forge: Efficient Dataset Condensation for Object Detection
  37. FreeMotion: A Unified Framework for Number-Free Text-to-Motion Synthesis
  38. LLaVA-VSD: Large Language-and-Vision Assistant for Visual Spatial Description
  39. MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation
  40. MambaAD: Exploring State Space Models for Multi-class Unsupervised Anomaly Detection
  41. MambaGesture: Enhancing Co-Speech Gesture Generation with Mamba and Disentangled Multi-Modality Fusion
  42. MotionBooth: Motion-Aware Customized Text-to-Video Generation
  43. PortraitBooth: A Versatile Portrait Model for Fast Identity-Preserved Personalization
  44. Real-IAD: A Real-World Multi-View Dataset for Benchmarking Versatile Industrial Anomaly Detection
  45. Rethinking Reverse Distillation for Multi-Modal Anomaly Detection
  46. Self-Supervised Likelihood Estimation with Energy Guidance for Anomaly Segmentation in Urban Scenes
  47. Self-supervised Feature Adaptation for 3D Industrial Anomaly Detection
  48. SuperSVG: Superpixel-Based Scalable Vector Graphics Synthesis
  49. TexDreamer: Towards Zero-Shot High-Fidelity 3D Human Texture Generation
  50. Towards Language-Driven Video Inpainting via Multimodal Large Language Models
  51. UniM-OV3D: Uni-Modality Open-Vocabulary 3D Scene Understanding with Fine-Grained Feature Representation