PPaperPicks

Rui Zhao

Sense Time Research, China

32 papers at tracked venues · 21 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. AIM 2025 Challenge on High FPS Motion Deblurring: Methods and Results
  2. Can an Individual Manipulate the Collective Decisions of Multi-Agents?
  3. DoraCycle: Domain-Oriented Adaptation of Unified Generative Model in Multimodal Cycles
    CVPR 2025 · Rui Zhao
  4. Efficient Multivariate Time Series Forecasting via Calibrated Language Models with Privileged Knowledge Distillation
  5. KITS: Inductive Spatio-Temporal Kriging with Increment Training Strategy
  6. On the Suitability of Reinforcement Fine-Tuning to Visual Tasks
  7. PUMA: Empowering Unified MLLM with Multi-Granular Visual Generation
  8. Re-Aligning Language to Visual Objects with an Agentic Workflow
  9. STAR: Efficient Preference-based Reinforcement Learning via Dual Regularization
  10. TimeCMA: Towards LLM-Empowered Multivariate Time Series Forecasting via Cross-Modality Alignment
  11. Towards Cross-Modality Modeling for Time Series Analytics: A Survey in the LLM Era
  12. Unlocking the Power of SAM 2 for Few-Shot Segmentation
  13. CLEAR: Can Language Models Really Understand Causal Graphs?
  14. CoSLight: Co-optimizing Collaborator Selection and Decision-making to Enhance Traffic Signal Control
  15. DragAnything: Motion Control for Anything Using Entity Representation
  16. DuaLight: Enhancing Traffic Signal Control by Leveraging Scenario-Specific and Scenario-Shared Knowledge
  17. DynVideo-E: Harnessing Dynamic NeRF for Large-Scale Motion- and View-Change Human-Centric Video Editing
  18. Eliminating Feature Ambiguity for Few-Shot Segmentation
  19. EvolveDirector: Approaching Advanced Text-to-Image Generation with Large Vision-Language Models
    NeurIPS 2024 · Rui Zhao
  20. Hybrid Mamba for Few-Shot Segmentation
  21. Instruct-ReID: A Multi-Purpose Person Re-Identification Task with Instructions
  22. InstructDET: Diversifying Referring Object Detection with Generalized Instructions
  23. MotionDirector: Motion Customization of Text-to-Video Diffusion Models
    ECCV 2024 · Rui Zhao
  24. Non-Neighbors Also Matter to Kriging: A New Contrastive-Prototypical Learning
  25. PDiT: Interleaving Perception and Decision-making Transformers for Deep Reinforcement Learning
  26. Self-Supervised Representation Learning from Arbitrary Scenarios
  27. Sequential Asynchronous Action Coordination in Multi-Agent Systems: A Stackelberg Decision Transformer Approach
  28. Sparse Global Matching for Video Frame Interpolation with Large Motion
  29. TPTU-v2: Boosting Task Planning and Tool Usage of Large Language Model-based Agents in Real-world Industry Systems
  30. VideoSwap: Customized Video Subject Swapping with Interactive Semantic Point Correspondence
  31. X- Adapter: Universal Compatibility of Plugins for Upgraded Diffusion Model
  32. X-Light: Cross-City Traffic Signal Control Using Transformer on Transformer as Meta Multi-Agent Reinforcement Learner