PPaperPicks

Yan Xia

12 papers at tracked venues · 9 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. CMMCoT: Enhancing Complex Multi-Image Comprehension via Multi-Modal Chain-of-Thought and Memory Augmentation
  2. AnomalyCoT: A Multi-Scenario Chain-of-Thought Dataset for Multimodal Large Language Models
  3. Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations
  4. CART: A Generative Cross-Modal Retrieval Framework With Coarse-To-Fine Semantic Modeling
  5. EAGER-LLM: Enhancing Large Language Models as Recommenders through Exogenous Behavior-Semantic Integration
  6. Enhancing Multimodal Unified Representations for Cross Modal Generalization
  7. Open-Set Cross Modal Generalization via Multimodal Unified Representation
  8. Overcoming both Domain Shift and Label Shift for Referring Video Segmentation
  9. RecBase: Generative Foundation Model Pretraining for Zero-Shot Recommendation
  10. Vela: Scalable Embeddings with Voice Large Language Models for Multimodal Retrieval
  11. EAGER: Two-Stream Generative Recommender with Behavior-Semantic Collaboration
  12. StyleSinger: Style Transfer for Out-of-Domain Singing Voice Synthesis