PPaperPicks

Zhenhui Ye

10 papers at tracked venues · 10 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. T2A-Feedback: Improving Basic Capabilities of Text-to-Audio Generation via Fine-grained AI Feedback
  2. AudioGPT: Understanding and Generating Speech, Music, Sound, and Talking Head
  3. Extending Multi-modal Contrastive Representations
  4. FreeBind: Free Lunch in Unified Multimodal Space via Knowledge Fusion
  5. InstructSpeech: Following Speech Editing Instructions via Large Language Models
  6. Make-A-Voice: Revisiting Voice Large Language Models as Scalable Multilingual and Multitask Learners
  7. Mega-TTS 2: Boosting Prompting Mechanisms for Zero-Shot Speech Synthesis
  8. MimicTalk: Mimicking a personalized and expressive 3D talking face in minutes
    NeurIPS 2024 · Zhenhui Ye
  9. Real3D-Portrait: One-shot Realistic 3D Talking Portrait Synthesis
    ICLR 2024 · Zhenhui Ye
  10. VoiceTuner: Self-Supervised Pre-training and Efficient Fine-tuning For Voice Generation