PPaperPicks

Fei Richard Yu

Carleton University, Department of Systems and Computer Engineering, Ottawa, ON, Canada

31 papers at tracked venues · 16 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. ForeDiffusion: Foresight-Conditioned Diffusion Policy via Future View Construction for Robot Manipulation
  2. M3SR: Multi-Scale Multi-Perceptual Mamba for Efficient Spectral Reconstruction
  3. PSPO: Prompt-Level Prioritization and Experience-Weighted Smoothing for Efficient Policy Optimization
  4. A Two-Stage Lightweight Framework for Efficient Land-Air Bimodal Robot Autonomous Navigation
  5. ABM++: Learning Generalizable Manipulation Policies with a Mask-Guided World Model
  6. Bi-directional Cable-driven Ankle Exoskeleton Coupled with Series Elastic Actuator for Compliant Gait Assisting*
  7. DSTR: Dual Scenes Transformer for Cross-Modal Fusion in 3D Object Detection
  8. EventGPT: Event Stream Understanding with Multimodal Large Language Models
  9. Ferret: Federated Full-Parameter Tuning at Scale for Large Language Models
  10. GaussianPU: Color Point Cloud Upsampling via 3D Gaussian Splatting
  11. Inter3D: A Benchmark and Strong Baseline for Human-Interactive 3D Object Reconstruction
  12. JAM: Keypoint-Guided Joint Prediction after Classification-Aware Marginal Proposal for Multi-Agent Interaction
  13. OnlineHOI: Towards Online Human-Object Interaction Generation and Perception
  14. PAFT: Prompt-Agnostic Fine-Tuning
  15. PointTalk: Audio-Driven Dynamic Lip Point Cloud for 3D Gaussian-based Talking Head Synthesis
  16. ReDit: Reward Dithering for Improved LLM Policy Optimization
  17. ReMask-Animate: Refined Character Image Animation Using Mask-Guided Adapters
  18. RoMa: A Robust Model Watermarking Scheme for Protecting IP in Diffusion Models
  19. Safe-Sora: Safe Text-to-Video Generation via Graphical Watermarking
  20. Subgraph Invariant Learning Towards Large-Scale Graph Node Classification
  21. TagGuideBot: Enhancing Robot Intelligence with Object Tags and VLMs
  22. Universal Visuo-Tactile Video Understanding for Embodied Interaction
  23. VideoHumanMIB: Unlocking Appearance Decoupling for Video Human Motion In-betweening
  24. WMarkGPT: Watermarked Image Understanding via Multimodal Large Language Models
  25. A Language-Driven Navigation Strategy Integrating Semantic Maps and Large Language Models
  26. ABM: Attention before Manipulation
  27. CodeSwap: Symmetrically Face Swapping Based on Prior Codebook
  28. LLaKey: Follow My Basic Action Instructions to Your Next Key State
  29. OTOcc: Optimal Transport for Occupancy Prediction
  30. OptEx: Expediting First-Order Optimization with Approximately Parallelized Iterations
  31. PP-TIL: Personalized Planning for Autonomous Driving with Instance-based Transfer Imitation Learning