PPaperPicks

Jianwei Yu

12 papers at tracked venues · 8 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. CoVoMix2: Advancing Zero-Shot Dialogue Generation with Fully Non-Autoregressive Flow Matching
  2. MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
  3. MoonCast: High-Quality Zero-Shot Podcast Generation
  4. MuCodec: Ultra Low-Bitrate Music Codec for Music Generation
  5. Pseudo-Autoregressive Neural Codec Language Models for Efficient Zero-Shot Text-to-Speech Synthesis
  6. SongBloom: Coherent Song Generation via Interleaved Autoregressive Sketching and Diffusion Refinement
  7. SongEditor: Adapting Zero-Shot Song Generation Language Model as a Multi-Task Editor
  8. WAKE: Watermarking Audio with Key Enrichment
  9. Comparing Discrete and Continuous Space LLMs for Speech Recognition
  10. Improved Factorized Neural Transducer Model For Text-only Domain Adaptation
  11. SECap: Speech Emotion Captioning with Large Language Model
  12. WeSep: A Scalable and Flexible Toolkit Towards Generalizable Target Speaker Extraction