PPaperPicks

Wei-Shi Zheng

Sun Yat-sen University, School of Information Science and Technology

62 papers at tracked venues · 50 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. DCAC: Dynamic Class-Aware Cache Creates Stronger Out-of-Distribution Detectors
  2. Exploring Surround-View Fisheye Camera 3D Object Detection
  3. TechCoach: Towards Technical-Point-Aware Descriptive Action Coaching
  4. AffordDexGrasp: Open-Set Language-Guided Dexterous Grasp With Generalizable-Instructive Affordance
  5. CLIP-RestoreX: Restore Image Structure and Perception in Exposure Correction
  6. Chain of Methodologies: Scaling Test Time Computation without Training
  7. ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction Generation
  8. DNF-Intrinsic: Deterministic Noise-Free Diffusion for Indoor Inverse Rendering
  9. Decoupled Distillation to Erase: A General Unlearning Method for Any Class-centric Tasks
  10. Diffusion-based Event Generation for High-Quality Image Deblurring
  11. Distilling LLM Prior to Flow Model for Generalizable Agent's Imagination in Object Goal Navigation
  12. EntityErasure: Erasing Entity Cleanly via Amodal Entity Segmentation and Completion
  13. FA: Forced Prompt Learning of Vision-Language Models for Out-of-Distribution Detection
  14. Global and Local Vision-Language Alignment for Few-Shot Learning and Few-Shot OOD Detection
  15. Hierarchical Vision-Language Learning for Medical Out-of-Distribution Detection
  16. LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of Large Language Models
  17. Learning Implicit Features with Flow-Infused Transformations for Realistic Virtual Try-On
  18. Less Static, More Private: Towards Transferable Privacy-Preserving Action Recognition by Generative Decoupled Learning
  19. Light-T2M: A Lightweight and Fast Model for Text-to-motion Generation
  20. MaintaAvatar: A Maintainable Avatar Based on Neural Radiance Fields by Continual Learning
  21. Modeling Multiple Normal Action Representations for Error Detection in Procedural Tasks
  22. Panorama Generation From NFoV Image Done Right
  23. ParGo: Bridging Vision-Language with Partial and Global Views
  24. Person De-reidentification: A Variation-guided Identity Shift Modeling
  25. ReferDINO: Referring Video Object Segmentation with Visual Grounding Foundations
  26. Rethinking Bimanual Robotic Manipulation: Learning with Decoupled Interaction Framework
  27. RoGSplat: Learning Robust Generalizable Human Gaussian Splatting from Sparse Multi-View Images
  28. Structure-Guided Diffusion Models for High-Fidelity Portrait Shadow Removal
  29. TacCap: A Wearable FBG-Based Tactile Sensor for Efficient Human-to-Robot Skill Transfer
  30. Task-Oriented 6-DoF Grasp Pose Detection in Clutters
  31. ViSpeak: Visual Instruction Feedback in Streaming Videos
  32. Viperson: Flexibly Generating Virtual Identity for Person Re-Identification
  33. When Shadow Removal Meets Intrinsic Image Decomposition: A Joint Learning Framework Using Unpaired Data
  34. iManip: Skill-Incremental Learning for Robotic Manipulation
  35. monoVLN: Bridging the Observation Gap between Monocular and Panoramic Vision and Language Navigation
  36. An Economic Framework for 6-DoF Grasp Detection
  37. Bridge Past and Future: Overcoming Information Asymmetry in Incremental Object Detection
  38. Deep Model Reference: Simple Yet Effective Confidence Estimation for Image Classification
  39. Dexterous Grasp Transformer
  40. DreamView: Injecting View-Specific Text Guidance Into Text-to-3D Generation
  41. Efficient and Effective Weakly-Supervised Action Segmentation via Action-Transition-Aware Boundary Alignment
  42. EgoExo-Fitness: Towards Egocentric and Exocentric Full-Body Action Understanding
  43. Exploiting Discrepancy in Feature Statistic for Out-of-Distribution Detection
  44. Factorized Diffusion Autoencoder for Unsupervised Disentangled Representation Learning
  45. FeatWalk: Enhancing Few-Shot Classification through Local View Leveraging
  46. FodFoM: Fake Outlier Data by Foundation Models Creates Stronger Visual Out-of-Distribution Detector
  47. Frozen-DETR: Enhancing DETR with Image Understanding from Frozen Foundation Models
  48. Grasp as You Say: Language-guided Dexterous Grasp Generation
  49. Loc4Plan: Locating Before Planning for Outdoor Vision and Language Navigation
  50. NECA: Neural Customizable Human Avatar
  51. PRET: Planning with Directed Fidelity Trajectory for Vision and Language Navigation
  52. PixelFade: Privacy-preserving Person Re-identification with Noise-guided Progressive Replacement
  53. Ranking Distillation for Open-Ended Video Question Answering with Insufficient Labels
  54. Rethinking Few-Shot Class-Incremental Learning: Learning from Yourself
  55. Revealing Distribution Discrepancy by Sampling Transfer in Unlabeled Data
  56. Sculpting Holistic 3D Representation in Contrastive Language-Image-3D Pre-Training
  57. Selective Hourglass Mapping for Universal Image Restoration Based on Diffusion Model
  58. SiFT: A Serial Framework with Textual Guidance for Federated Learning
  59. Siamese Learning with Joint Alignment and Regression for Weakly-Supervised Video Paragraph Grounding
  60. Single-View Scene Point Cloud Human Grasp Generation
  61. SynopGround: A Large-Scale Dataset for Multi-Paragraph Video Grounding from TV Dramas and Synopses
  62. TagFog: Textual Anchor Guidance and Fake Outlier Generation for Visual Out-of-Distribution Detection