PPaperPicks

Zhongmin Cai

10 papers at tracked venues · 9 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. A Multi-Stage Structural Captioning Framework for Enhancing Chinese Image-to-Video Generation in Baidu
  2. AReframedChair: Reframing the Empty Chair through Dyadic and Triadic AR-Mediated Self-Embodiment
  3. DroidRetriever: A Transparent and Steerable Automation System for Collaborative Mobile Information Seeking
  4. EyeSee: Enhancing Art Appreciation through Anthropomorphic Interpretations from Multiple Perspectives
  5. Predicting User Behavior in Smart Spaces with LLM-Enhanced Logs and Personalized Prompts
  6. Retrieval-Augmented Image Captioning and Generation with Entity Concepts Enhancement for Baidu Multimodal Advertising
  7. Beyond Single Stationary Policies: Meta-Task Players as Naturally Superior Collaborators
  8. QGEval: Benchmarking Multi-dimensional Evaluation for Question Generation
  9. Towards Building Condition-Based Cross-Modality Intention-Aware Human-AI Cooperation under VR Environment
  10. VisionTasker: Mobile Task Automation Using Vision Based UI Understanding and LLM Task Planning