PPaperPicks

Yang Liu

Peking University, Wangxuan Institute of Computer Technology, Beijing, China

28 papers at tracked venues · 22 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. 3DWG: 3D Weakly Supervised Visual Grounding via Category and Instance-Level Alignment
  2. AR-VRM: Imitating Human Motions for Visual Robot Manipulation with Analogical Reasoning
  3. Advancing 3D Scene Understanding with MV-ScanQA Multi-View Reasoning Evaluation and TripAlign Pre-training Dataset
  4. Balancing Preservation and Modification: A Region and Semantic Aware Metric for Instruction-Based Image Editing
  5. ConMo: Controllable Motion Disentanglement and Recomposition for Zero-Shot Motion Transfer
  6. Generative Video Diffusion for Unseen Novel Semantic Video Moment Retrieval
  7. Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement
  8. Interact-Custom: Customized Human Object Interaction Image Generation
  9. InteractMove: Text-Controlled Human-Object Interaction Generation in 3D Scenes with Movable Objects
  10. Investigating Domain Gaps for Indoor 3D Object Detection
  11. Open-Vocabulary Hoi Detection With Interaction-Aware Prompt and Concept Calibration
  12. PlanLLM: Video Procedure Planning with Refinable Large Language Models
  13. Socratic Style Chain-of-Thoughts Help LLMs to be a Better Reasoner
  14. TRKT: Weakly Supervised Dynamic Scene Graph Generation with Temporal-Enhanced Relation-Aware Knowledge Transferring
  15. Zero Shot Domain Adaptive Semantic Segmentation by Synthetic Data Generation and Progressive Adaptation
  16. 3D Vision and Language Pretraining with Large-Scale Synthetic Data
  17. Active Object Detection with Knowledge Aggregation and Distillation from Large Models
  18. Bridging the Gap between 2D and 3D Visual Question Answering: A Fusion Approach for 3D VQA
  19. Exploring Conditional Multi-modal Prompts for Zero-Shot HOI Detection
  20. Exploring the Potential of Large Foundation Models for Open-Vocabulary HOI Detection
  21. Novel Class Discovery in Chest X-rays via Paired Images and Text
  22. OED: Towards One-stage End-to-End Dynamic Scene Graph Generation
  23. RelScene: A Benchmark and baseline for Spatial Relations in text-driven 3D Scene Generation
  24. ResVG: Enhancing Relation and Semantic Understanding in Multiple Instances for Visual Grounding
  25. Semantic-Aware Human Object Interaction Image Generation
  26. Semantic-Guided Novel Category Discovery
  27. Training-Free Video Temporal Grounding Using Large-Scale Pre-trained Models
  28. WAS: Dataset and Methods for Artistic Text Segmentation