PPaperPicks

Zhuo Chen

15 papers at tracked venues · 11 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Sparse Causal Latent Features for Robust Multimodal Learning under Distribution Shifts
  2. Advancing Zero-shot Text-to-Speech Intelligibility across Diverse Domains via Preference Alignment
  3. Compact R-X-Y Stage and Dual-Finger Micromanipulator under Inverted Optical Microscope for Microassembly
  4. DiTAR: Diffusion Transformer Autoregressive Modeling for Speech Generation
  5. ELLA-V: Stable Neural Codec Language Modeling with Alignment-Guided Sequence Reordering
  6. Language Model Can Listen While Speaking
  7. MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
  8. SilentStriker: Toward Stealthy Bit-Flip Attacks on Large Language Models
  9. Sounding that Object: Interactive Object-Aware Image to Audio Generation
  10. Theoretical Insights in Model Inversion Robustness and Conditional Entropy Maximization for Collaborative Inference Systems
  11. Towards Reliable Large Audio Language Model
  12. Which Tasks Should Be Compressed Together? A Causal Discovery Approach for Efficient Multi-Task Representation Compression
  13. A Unified Image Compression Method for Human Perception and Multiple Vision Tasks
  14. COSMIC: Data Efficient Instruction-tuning For Speech In-Context Learning
  15. TacoLM: GaTed Attention Equipped Codec Language Model are Efficient Zero-Shot Text to Speech Synthesizers