PPaperPicks

Paul Hongsuck Seo

16 papers at tracked venues · 10 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. GOAT: A Training Framework for Goal-Oriented Agent with Tools
  2. Bridging Audio and Vision: Zero-Shot Audiovisual Segmentation by Connecting Pretrained Models
  3. Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries
  4. DGMO: Training-Free Audio Source Separation through Diffusion-Guided Mask Optimization
  5. DialNav: Multi-Turn Dialog Navigation with a Remote Guide
  6. LCIRC: A Recurrent Compression Approach for Efficient Long-form Context and Query Dependent Modeling in LLMs
  7. Multi-Granularity Video Object Segmentation
  8. Random Conditioning for Diffusion Model Compression with Distillation
  9. ReSCORE: Label-free Iterative Retriever Training for Multi-hop Question Answering with Relevance-Consistency Supervision
  10. ReTAG: Retrieval-Enhanced, Topic-Augmented Graph-Based Global Sensemaking
  11. Seg4Diff: Unveiling Open-Vocabulary Semantic Segmentation in Text-to-Image Diffusion Transformers
  12. CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation
  13. Learning Correlation Structures for Vision Transformers
  14. Pseudo-RIS: Distinctive Pseudo-Supervision Generation for Referring Image Segmentation
  15. Towards Open-Vocabulary Semantic Segmentation Without Semantic Labels
  16. TrackIME: Enhanced Video Point Tracking via Instance Motion Estimation