PPaperPicks

Wen-Huang Cheng

National Chiao Tung University, Taiwan

30 papers at tracked venues · 24 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. DirectDrag: High-Fidelity, Mask-Free, Prompt-Free Drag-based Image Editing via Readout-Guided Feature Alignment
  2. Dragonite: Single-Step Drag-based Image Editing with Geometric-Semantic Guidance
  3. CookAnything: A Framework for Flexible and Consistent Multi-Step Recipe Image Generation
  4. EmoArt: A Multidimensional Dataset for Emotion-Aware Artistic Generation
  5. Flowing Crowd to Count Flows: A Self-Supervised Framework for Video Individual Counting
  6. From Prompt to Progression: Taming Video Diffusion Models for Seamless Attribute Transition
  7. Future Sight and Tough Fights: Revolutionizing Sequential Recommendation with FENRec
  8. InstructFLIP: Exploring Unified Vision-Language Model for Face Anti-spoofing
  9. MEGC2025: Micro-Expression Grand Challenge on Spot Then Recognize and Visual Question Answering
  10. Memory-Augmented Re-Completion for 3D Semantic Scene Completion
  11. MonoTAKD: Teaching Assistant Knowledge Distillation for Monocular 3D Object Detection
  12. OinkTrack: An Ultra-Long-Term Dataset for Multi-Object Tracking and Re-Identification of Group-Housed Pigs
  13. Personalized Lip Reading: Adapting to Your Unique Lip Movements with Vision and Language
  14. Perspective-Aware Teaching: Adapting Knowledge for Heterogeneous Distillation
  15. Proceedings of the 33rd ACM International Conference on Multimedia, MM 2025, Dublin, Ireland, October 27-31, 2025
  16. Radiance Field-Based Pose Estimation via Decoupled Optimization Under Challenging Initial Conditions
  17. RecipeGen: A Step-Aligned Multimodal Benchmark for Real-World Recipe Generation
  18. Relightable and Dynamic Gaussian Avatar Reconstruction from Monocular Video
  19. SMPV: Social Media Prediction for Videos
  20. Training-Free Industrial Defect Generation with Diffusion Models
  21. When Anchors Meet Cold Diffusion: A Multi-Stage Approach to Lane Detection
  22. DQ-DETR: DETR with Dynamic Query for Tiny Object Detection
  23. Distraction is All You Need: Memory-Efficient Image Immunization against Diffusion-Based Image Editing
  24. EmoVIT: Revolutionizing Emotion Insights with Visual Instruction Tuning
  25. MEGC2024: ACM Multimedia 2024 Facial Micro-Expression Grand Challenge
  26. ReCorD: Reasoning and Correcting Diffusion for HOI Generation
  27. SMP Challenge Summary: Social Media Prediction Challenge
  28. The Fabrication of Reality and Fantasy: Scene Generation with LLM-Assisted Prompt Interpretation
  29. TrajFine: Predicted Trajectory Refinement for Pedestrian Trajectory Forecasting
  30. TrajPrompt: Aligning Color Trajectory with Vision-Language Representations