PPaperPicks

Yuankai Qi

25 papers at tracked venues · 21 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. ImVision: Adapting Pretrained Vision Models for Time-Series Imputation
  2. InstructDubber: Instruction-based Alignment for Zero-shot Movie Dubbing
  3. The Devil is in the Distributions: Explicit Modeling of Scene Content is Key in Zero-Shot Video Captioning
  4. Tracking the Unstable: Appearance-Guided Motion Modeling for Robust Multi-Object Tracking in UAV-Captured Videos
  5. CausalMVC: Causal Content-Style Representation Learning for Deep Multi-View Clustering
  6. EmoDubber: Towards High Quality and Emotion Controllable Movie Dubbing
  7. FlowDubber: Movie Dubbing with LLM-based Semantic-aware Learning and Flow Matching based Voice Enhancing
  8. Generating Synthetic Data for Unsupervised Federated Learning of Cross-Modal Retrieval
  9. Incomplete Multi-View Multi-Label Classification via Diffusion-Guided Redundancy Removal
  10. Learning from Uncertainty: A Cloud-Based Active Learning for Detecting Bone Union in Mandibular Reconstruction
  11. Medusa: A Multi-Scale High-order Contrastive Dual-Diffusion Approach for Multi-View Clustering
  12. Prosody-Enhanced Acoustic Pre-training and Acoustic-Disentangled Prosody Adapting for Movie Dubbing
  13. SDVPT: Semantic-Driven Visual Prompt Tuning for Open-world Object Counting
  14. STGS: Spatio-temporal Graph Sparsification Using Reinforcement Learning
  15. Seeing the Trees for the Forest: Rethinking Weakly-Supervised Medical Visual Grounding
  16. Separation of Powers: On Segregating Knowledge from Observation in LLM-enabled Knowledge-based Visual Question Answering
  17. Trffc: Efficient Traffic Forecasting through Adaptive Spatio-Temporal Graph Reduction
  18. Visual and Semantic Prompt Collaboration for Generalized Zero-Shot Learning
  19. Augmented Commonsense Knowledge for Remote Object Grounding
  20. Decomposing Disease Descriptions for Enhanced Pathology Detection: A Multi-Aspect Vision-Language Pre-Training Framework
  21. From Speaker to Dubber: Movie Dubbing with Prosody and Duration Consistency Learning
  22. Generating Content for HDR Deghosting from Frequency View
  23. Structural Attention: Rethinking Transformer for Unpaired Medical Image Synthesis
  24. StyleDubber: Towards Multi-Scale Style Learning for Movie Dubbing
  25. Weakly Supervised Video Individual Counting