PPaperPicks

Yufan Zhou

11 papers at tracked venues · 8 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. ARTIST: Improving the Generation of Text-Rich Images with Disentangled Diffusion Models and Large Language Models
  2. Grounded-VideoLLM: Sharpening Fine-grained Temporal Grounding in Video Large Language Models
  3. Multimodal LLMs as Customized Reward Models for Text-to-Image Generation
  4. Numerical Pruning for Efficient Autoregressive Models
  5. SV-RAG: LoRA-Contextualizing Adaptation of MLLMs for Long Document Understanding
  6. TTVD: Towards a Geometric Framework for Test-Time Adaptation Based on Voronoi Diagram
  7. Customization Assistant for Text-to-image Generation
    CVPR 2024 · Yufan Zhou
  8. Navigating the Dual Facets: A Comprehensive Evaluation of Sequential Memory Editing in Large Language Models
  9. TRINS: Towards Multimodal Language Models that Can Read
  10. TextLap: Customizing Language Models for Text-to-Layout Planning
  11. Towards Aligned Layout Generation via Diffusion Model with Aesthetic Constraints