PPaperPicks

Simon Jenni

10 papers at tracked venues · 8 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. Improving Large Vision and Language Models by Learning from a Panel of Peers
  2. MAGNET: Augmenting Generative Decoders with Representation Learning and Infilling Capabilities
  3. The Indra Representation Hypothesis for Multimodal Alignment
  4. The Photographer's Eye: Teaching Multimodal Large Language Models to See, and Critique Like Photographers
  5. ViDROP: Video Dense Representation through Spatio-Temporal Sparsity
  6. Building Vision-Language Models on Solid Foundations with Masked Distillation
  7. Concept Weaver: Enabling Multi-Concept Fusion in Text-to-Image Models
  8. FineMatch: Aspect-Based Fine-Grained Image and Text Mismatch Detection and Correction
  9. No More Shortcuts: Realizing the Potential of Temporal Self-Supervision
    AAAI 2024 ·
    Ishan Rajendrakumar Dave
  10. Sync from the Sea: Retrieving Alignable Videos from Large-Scale Datasets
    ECCV 2024 ·
    Ishan Rajendrakumar Dave