P
PaperPicks
Conferences
Xianwei Zhuang
23 papers at tracked venues · 12 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
ACL
×6
EMNLP
×4
AAAI
×3
InterSpeech
×3
ECCV
×2
ACM MM
×1
CVPR
×1
ICLR
×1
IJCAI
×1
NAACL
×1
Frequent coauthors
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
×3
Xuxin Cheng
DBLP profile ↗
ORCID search ↗
×3
Liming Liang
DBLP profile ↗
ORCID search ↗
×2
Zhanpeng Chen
DBLP profile ↗
ORCID search ↗
×2
Lexiang Tang
DBLP profile ↗
ORCID search ↗
×1
Yuguo Yin
DBLP profile ↗
ORCID search ↗
×1
Yuxin Xie
DBLP profile ↗
ORCID search ↗
×1
Xuan Wu
DBLP profile ↗
ORCID search ↗
×1
Wanshi Xu
DBLP profile ↗
ORCID search ↗
×1
Papers
Not All Tokens and Heads Are Equally Important: Dual-Level Attention Intervention for Hallucination Mitigation
AAAI 2026
·
Lexiang Tang
DBLP profile ↗
ORCID search ↗
ATRI: Mitigating Multilingual Audio Text Retrieval Inconsistencies by Reducing Data Distribution Errors
ACL 2025
·
Yuguo Yin
DBLP profile ↗
ORCID search ↗
Can We Trust AI Doctors? A Survey of Medical Hallucination in Large Language and Large Vision-Language Models
ACL 2025
·
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
FoleyMaster: High-Quality Video-to-Audio Synthesis via MLLM-Augmented Prompt Tuning and Joint Semantic-Temporal Adaptation
InterSpeech 2025
·
Liming Liang
DBLP profile ↗
ORCID search ↗
SpeechSEC: A Unified Multi-Task Framework for Speech Synthesis, Editing, and Continuation
InterSpeech 2025
·
Liming Liang
DBLP profile ↗
ORCID search ↗
UniCoTT: A Unified Framework for Structural Chain-of-Thought Distillation
ICLR 2025
·
Xianwei Zhuang
VASparse: Towards Efficient Visual Hallucination Mitigation via Visual-Aware Token Sparsification
CVPR 2025
·
Xianwei Zhuang
Code-Switching Can be Better Aligners: Advancing Cross-Lingual SLU through Representation-Level and Prediction-Level Alignment
ACL 2024
·
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
Cyclical Contrastive Learning Based on Geodesic for Zero-shot Cross-lingual Spoken Language Understanding
ACL 2024
·
Xuxin Cheng
DBLP profile ↗
ORCID search ↗
Dual-oriented Disentangled Network with Counterfactual Intervention for Multimodal Intent Detection
EMNLP 2024
·
Zhanpeng Chen
DBLP profile ↗
ORCID search ↗
GPA: Global and Prototype Alignment for Audio-Text Retrieval
InterSpeech 2024
·
Yuxin Xie
DBLP profile ↗
ORCID search ↗
Game on Tree: Visual Hallucination Mitigation via Coarse-to-Fine View Tree and Game Theory
EMNLP 2024
·
Xianwei Zhuang
KDProR: A Knowledge-Decoupling Probabilistic Framework for Video-Text Retrieval
ECCV 2024
·
Xianwei Zhuang
MaCSC: Towards Multimodal-augmented Pre-trained Language Models via Conceptual Prototypes and Self-balancing Calibration
NAACL 2024
·
Xianwei Zhuang
MoE-SLU: Towards ASR-Robust Spoken Language Understanding via Mixture-of-Experts
ACL 2024
·
Xuxin Cheng
DBLP profile ↗
ORCID search ↗
PCAD: Towards ASR-Robust Spoken Language Understanding via Prototype Calibration and Asymmetric Decoupling
ACL 2024
·
Xianwei Zhuang
Relevance Is a Guiding Light: Relevance-aware Adaptive Learning for End-to-end Task-oriented Dialogue System
EMNLP 2024
·
Zhanpeng Chen
DBLP profile ↗
ORCID search ↗
TFCD: Towards Multi-modal Sarcasm Detection via Training-Free Counterfactual Debiasing
IJCAI 2024
·
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
Towards Explainable Joint Models via Information Theory for Multiple Intent Detection and Slot Filling
AAAI 2024
·
Xianwei Zhuang
Towards Multi-Intent Spoken Language Understanding via Hierarchical Attention and Optimal Transport
AAAI 2024
·
Xuxin Cheng
DBLP profile ↗
ORCID search ↗
Towards Multimodal-augmented Pre-trained Language Models via Self-balanced Expectation-Maximization Iteration
ACM MM 2024
·
Xianwei Zhuang
Uncertainty-Aware Sign Language Video Retrieval with Probability Distribution Modeling
ECCV 2024
·
Xuan Wu
DBLP profile ↗
ORCID search ↗
What are the Generator Preferences for End-to-end Task-Oriented Dialog System?
EMNLP 2024
·
Wanshi Xu
DBLP profile ↗
ORCID search ↗