PPaperPicks

Sungroh Yoon

39 papers at tracked venues · 26 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. CANDI: Curated Test-Time Adaptation for Multivariate Time-Series Anomaly Detection Under Distribution Shift
  2. DCText: Scheduled Attention Masking for Visual Text Generation via Divide-and-Conquer Strategy
  3. GDoFS: Gaussian DoF Separation for Plausible 3D Geometry in Sparse-View 3DGS
  4. Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
  5. SAVE: Sparse Autoencoder-Driven Visual Information Enhancement for Mitigating Object Hallucination
  6. Still Between Us? Evaluating and Improving Voice Assistant Robustness to Third-Party Interruptions
  7. Style-Friendly SNR Sampler for Style-Driven Generation
  8. Verbal-R3: Verbal Reranker as the Missing Bridge between Retrieval and Reasoning
  9. Battling the Non-stationarity in Time Series Forecasting via Test-time Adaptation
  10. Causality-Aware Contrastive Learning for Robust Multivariate Time-Series Anomaly Detection
  11. Correcting Negative Bias in Large Language Models through Negative Attention Score Alignment
  12. DefectFill: Realistic Defect Generation with Inpainting Diffusion Model for Visual Inspection
  13. Disentangled Motion Modeling for Video Frame Interpolation
  14. Does Your Voice Assistant Remember? Analyzing Conversational Context Recall and Utilization in Voice Interaction Models
  15. EdiText: Controllable Coarse-to-Fine Text Editing with Diffusion Language Models
  16. Exploring the Potential of LLMs as Personalized Assistants: Dataset, Evaluation, and Analysis
  17. Interpretable Next-token Prediction via the Generalized Induction Head
  18. Know "No" Better: A Data-Driven Approach for Enhancing Negation Awareness in CLIP
  19. Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator
  20. RePIC: Reinforced Post-Training for Personalizing Multi-Modal Language Models
  21. Rethinking Training for De-biasing Text-to-Image Generation: Unlocking the Potential of Stable Diffusion
  22. Toward Robust Hyper-Detailed Image Captioning: A Multiagent Approach and Dual Evaluation Metrics for Factuality and Coverage
  23. Unleashing Multi-Hop Reasoning Potential in Large Language Models through Repetition of Misordered Context
  24. Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models
  25. CKNN: Cleansed k-Nearest Neighbor for Unsupervised Video Anomaly Detection
  26. ControlDreamer: Blending Geometry and Style in Text-to-3D
  27. Controlled Text Generation for Black-box Language Models via Score-based Progressive Editor
  28. DAFA: Distance-Aware Fair Adversarial Training
  29. Efficient Diffusion-Driven Corruption Editor for Test-Time Adaptation
  30. Entropy is not Enough for Test-Time Adaptation: From the Perspective of Disentangled Factors
  31. Interactive Text-to-Image Retrieval with Large Language Models: A Plug-and-Play Approach
  32. Introducing Spectral Attention for Long-Range Dependency in Time Series Forecasting
  33. LLM-based Frameworks for API Argument Filling in Task-Oriented Conversational Systems
  34. Paralinguistics-Aware Speech-Empowered Large Language Models for Natural Conversation
  35. SF(DA)2: Source-free Domain Adaptation Through the Lens of Data Augmentation
  36. Semantic Token Reweighting for Interpretable and Controllable Text Embeddings in CLIP
  37. Textual Training for the Hassle-Free Removal of Unwanted Visual Data: Case Studies on OOD and Hateful Image Detection
  38. Unsupervised Homography Estimation on Multimodal Image Pair via Alternating Optimization
  39. VoiceTailor: Lightweight Plug-In Adapter for Diffusion-Based Personalized Text-to-Speech