PPaperPicks

Kyogu Lee

15 papers at tracked venues · 6 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. Few-step Adversarial Schrödinger Bridge for Generative Speech Enhancement
  2. MGE-LDM: Joint Latent Diffusion for Simultaneous Music Generation and Source Extraction
  3. Multidimensional Adaptive Coefficient for Inference Trajectory Optimization in Flow and Diffusion
  4. Song Form-aware Full-Song Text-to-Lyrics Generation with Multi-Level Granularity Syllable Count Control
  5. SynthRL: Cross-domain Synthesizer Sound Matching via Reinforcement Learning
  6. Towards Bitrate-Efficient and Noise-Robust Speech Coding with Variable Bitrate RVQ
  7. Uncertainty-Aware Self-Training for CTC-Based Automatic Speech Recognition
  8. Understanding Audio-Text Retrieval Through Singular Value Decomposition
  9. Vo-Ve: An Explainable Voice-Vector for Speaker Identity Evaluation
  10. Voice-Based Dysphagia Detection: Leveraging Self-Supervised Speech Representation
  11. Differentiable Modal Synthesis for Physical Modeling of Planar String Sound and Motion Simulation
  12. Distance Sampling-based Paraphraser Leveraging ChatGPT for Text Data Manipulation
  13. Emosical: An Emotion-Annotated Musical Theatre Dataset
  14. Guiding Frame-Level CTC Alignments Using Self-knowledge Distillation
  15. Hear Your Face: Face-based voice conversion with F0 estimation