PPaperPicks

Peng Jin

15 papers at tracked venues · 10 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Next Patch Prediction for AutoRegressive Visual Generation
  2. Aligning Instance Brownian Bridge with Texts for Open-Vocabulary Video Instance Segmentation
  3. LlaVA-CoT: Let Vision Language Models Reason Step-By-Step
  4. MUSE: Mamba Is Efficient Multi-scale Learner for Text-video Retrieval
  5. MoE++: Accelerating Mixture-of-Experts Methods with Zero-Computation Experts
    ICLR 2025 · Peng Jin
  6. MoH: Multi-Head Attention as Mixture-of-Head Attention
    ICML 2025 · Peng Jin
  7. Orthogonal Subspace Decomposition for Generalizable AI-Generated Image Detection
  8. Chat-UniVi: Unified Visual Representation Empowers Large Language Models with Image and Video Understanding
    CVPR 2024 · Peng Jin
  9. FreestyleRet: Retrieving Images from Style-Diversified Queries
  10. LOOK-M: Look-Once Optimization in KV Cache for Efficient Multimodal Long-Context Inference
  11. Local Action-Guided Motion Diffusion Model for Text-to-Motion Generation
    ECCV 2024 · Peng Jin
  12. Parallel Vertex Diffusion for Unified Visual Grounding
  13. RAP: Efficient Text-Video Retrieval with Sparse-and-Correlated Adapter
  14. Repaint123: Fast and High-Quality One Image to 3D Generation with Progressive Controllable Repainting
  15. Video-LLaVA: Learning United Visual Representation by Alignment Before Projection