PPaperPicks

Bin Zhu

8 papers at tracked venues · 6 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Next Patch Prediction for AutoRegressive Visual Generation
  2. Table-as-Search: Agentic Information Seeking is Table Completion
  3. DreamDance: Animating Human Images by Enriching 3D Geometry Cues from 2D Poses
  4. Hand1000: Generating Realistic Hands from Text with Only 1, 000 Images
  5. Multimodal Interpretable Depression Analysis Using Visual, Physiological, Audio and Textual Data
  6. PolarNeXt: Rethink Instance Segmentation with Polar Representation
  7. LanguageBind: Extending Video-Language Pretraining to N-modality by Language-based Semantic Alignment
    ICLR 2024 · Bin Zhu
  8. Video-LLaVA: Learning United Visual Representation by Alignment Before Projection