PPaperPicks

Haoyuan Li

12 papers at tracked venues · 12 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. CMMCoT: Enhancing Complex Multi-Image Comprehension via Multi-Modal Chain-of-Thought and Memory Augmentation
  2. MAU-GPT: Enhancing Multi-type Industrial Anomaly Understanding via Anomaly-aware and Generalist Experts Adaptation
  3. Align²LLaVA: Cascaded Human and Large Language Model Preference Alignment for Multi-modal Instruction Curation
  4. Detecting and Mitigating Hallucination in Large Vision Language Models via Fine-Grained AI Feedback
  5. HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation
  6. LLaVA-MoD: Making LLaVA Tiny via MoE-Knowledge Distillation
  7. MARS: Mixture of Auto-Regressive Models for Fine-grained Text-to-image Synthesis
  8. Streaming Video Question-Answering with In-context Video KV-Cache Retrieval
  9. T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts
  10. TeamLoRA: Boosting Low-Rank Adaptation with Expert Collaboration and Competition
  11. EAGER: Two-Stream Generative Recommender with Behavior-Semantic Collaboration
  12. T2S-GPT: Dynamic Vector Quantization for Autoregressive Sign Language Production from Text