PPaperPicks

Yan Wang

38 papers at tracked venues · 27 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Compressing then Matching: An Efficient Pre-training Paradigm for Multimodal Embedding
  2. DualFete: Revisiting Teacher-Student Interactions from a Feedback Perspective for Semi-supervised Medical Image Segmentation
  3. EchoMimicV3: 1.3B Parameters Are All You Need for Unified Multi-Modal and Multi-Task Human Animation
  4. How They Type: Eye and Finger Movement Strategies in Typing of Individuals with Cerebral Palsy
  5. LindormVector: A Distributed Vector Engine on a Cloud-Native Multi-Model NoSQL Database
    SIGMOD 2026 · Yan Wang
  6. RAM-SD: Retrieval-Augmented Multi-agent framework for Sarcasm Detection
  7. Semore: VLM-guided Enhanced Semantic Motion Representations for Visual Reinforcement Learning
  8. SparseWorld: A Flexible, Adaptive, and Efficient 4D Occupancy World Model Powered by Sparse and Dynamic Queries
  9. Block-Attention for Efficient Prefilling
  10. F-Adapter: Frequency-Adaptive Parameter-Efficient Fine-Tuning in Scientific Machine Learning
  11. GapMatch: Bridging Instance and Model Perturbations for Enhanced Semi-Supervised Medical Image Segmentation
  12. Integrating Task-Specific and Universal Adapters for Pre-Trained Model-Based Class-Incremental Learning
    ICCV 2025 · Yan Wang
  13. LEANCODE: Understanding Models Better for Code Simplification of Pre-trained Large Language Models
    ACL 2025 · Yan Wang
  14. MamV2XCalib: V2X-based Target-Less Infrastructure Camera Calibration with State Space Model
  15. MonoMixer: Marrying Convolution and Vision Transformer for Efficient Self-Supervised Monocular Depth Estimation
  16. NTIRE 2025 Challenge on Short-form UGC Video Quality Assessment and Enhancement: Methods and Results
  17. Omni-Fusion of Spatial and Spectral for Hyperspectral Image Segmentation
  18. Phys4DRT: Physics-based 4D Generation for Real-Time Interaction with Time-Frequency Supervision
  19. Physical-aware Neural Radiance Fields for Efficient Exposure Correction
  20. ProteinBench: A Holistic Evaluation of Protein Foundation Models
  21. Reactffusion: Physical Contact-guided Diffusion Model for Reaction Generation
  22. Semi-Supervised Vision-Centric 3D Occupancy World Model for Autonomous Driving
  23. Simultaneous Modeling of Protein Conformation and Dynamics via Autoregression
  24. Synthetic Dysarthric Speech: A Supplement, Not a Substitute for Authentic Data in Dysarthric Speech Recognition
  25. A User-Friendly Framework for Generating Model-Preferred Prompts in Text-to-Image Synthesis
  26. Automatic Captioning based on Visible and Infrared Images
    ICRA 2024 · Yan Wang
  27. Cliff: Leveraging Ambiguous Samples for Enhanced Test-Time Adaptation
  28. CoAtFormer: Vision Transformer with Composite Attention
  29. Instruction-Driven Game Engine: A Poker Case Study
  30. LUCIDA: Low-Dose Universal-Tissue CT Image Domain Adaptation for Medical Segmentation
  31. MLPER: Multi-Level Prompts for Adaptively Enhancing Vision-Language Emotion Recognition
  32. Probability-Polarized Optimal Transport for Unsupervised Domain Adaptation
    AAAI 2024 · Yan Wang
  33. Protein Conformation Generation via Force-Guided SE(3) Diffusion Models
    ICML 2024 · Yan Wang
  34. Real-World Visual Navigation for Cardiac Ultrasound View Planning
  35. SRFUND: A Multi-Granularity Hierarchical Structure Reconstruction Benchmark in Form Understanding
  36. The Ninth NTIRE 2024 Efficient Super-Resolution Challenge Report
  37. Unleashing the Potentials of Likelihood Composition for Multi-modal Language Models
  38. Visual-Augmented Dynamic Semantic Prototype for Generative Zero-Shot Learning