P
PaperPicks
Conferences
Xuxin Cheng
41 papers at tracked venues · 24 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
ACL
×10
AAAI
×6
ICLR
×4
ACM MM
×3
EMNLP
×3
ECCV
×2
ICRA
×2
InterSpeech
×2
CIKM
×1
CVPR
×1
EACL
×1
IJCAI
×1
IROS
×1
MICCAI
×1
NAACL
×1
RSS
×1
WSDM
×1
Frequent coauthors
Xianwei Zhuang
DBLP profile ↗
ORCID search ↗
×6
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
×6
Ziyu Yao
DBLP profile ↗
ORCID search ↗
×2
Hongxiang Li
DBLP profile ↗
ORCID search ↗
×2
Yifei Xin
DBLP profile ↗
ORCID search ↗
×2
Wanshi Xu
DBLP profile ↗
ORCID search ↗
×2
Jiaming Zhou
DBLP profile ↗
ORCID search ↗
×1
Yubo Jiang
DBLP profile ↗
ORCID search ↗
×1
Siwei Wu
DBLP profile ↗
ORCID search ↗
×1
Yuzhe Zhang
DBLP profile ↗
ORCID search ↗
×1
Xin Yang
DBLP profile ↗
ORCID search ↗
×1
Jingheng Ye
DBLP profile ↗
ORCID search ↗
×1
Papers
DIFFA-2: A Practical Diffusion Large Language Model for General Audio Understanding
ACL 2026
·
Jiaming Zhou
DBLP profile ↗
ORCID search ↗
Global Context or Local Detail? Adaptive Visual Grounding for Hallucination Mitigation
ACL 2026
·
Yubo Jiang
DBLP profile ↗
ORCID search ↗
MMRA: A Benchmark for Evaluating Multi-Granularity and Multi-Image Relational Association Capabilities in Large Visual Language Models
EACL 2026
·
Siwei Wu
DBLP profile ↗
ORCID search ↗
SILO-BENCH: A Scalable Environment for Evaluating Distributed Coordination in Multi-Agent LLM Systems
ACL 2026
·
Yuzhe Zhang
DBLP profile ↗
ORCID search ↗
When 20 Agents Fail to Sort: The Distributed Sorting Benchmark for Scalable Multi-Agent Systems
ACL 2026
·
Xin Yang
DBLP profile ↗
ORCID search ↗
CountLLM: Towards Generalizable Repetitive Action Counting via Large Language Model
CVPR 2025
·
Ziyu Yao
DBLP profile ↗
ORCID search ↗
DisPose: Disentangling Pose Guidance for Controllable Human Image Animation
ICLR 2025
·
Hongxiang Li
DBLP profile ↗
ORCID search ↗
EXCGEC: A Benchmark for Edit-Wise Explainable Chinese Grammatical Error Correction
AAAI 2025
·
Jingheng Ye
DBLP profile ↗
ORCID search ↗
Helpful DoggyBot: Open-World Object Fetching using Legged Robots and Vision-Language Models
IROS 2025
·
Qi Wu
DBLP profile ↗
ORCID search ↗
Mobile-TeleVision: Predictive Motion Priors for Humanoid Whole-Body Control
ICRA 2025
·
Chenhao Lu
DBLP profile ↗
ORCID search ↗
UniCoTT: A Unified Framework for Structural Chain-of-Thought Distillation
ICLR 2025
·
Xianwei Zhuang
DBLP profile ↗
ORCID search ↗
Aligner²: Enhancing Joint Multiple Intent Detection and Slot Filling via Adjustive and Forced Cross-Task Alignment
AAAI 2024
·
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
Audio-text Retrieval with Transformer-based Hierarchical Alignment and Disentangled Cross-modal Representation
InterSpeech 2024
·
Yifei Xin
DBLP profile ↗
ORCID search ↗
Code-Switching Can be Better Aligners: Advancing Cross-Lingual SLU through Representation-Level and Prediction-Level Alignment
ACL 2024
·
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
Cyclical Contrastive Learning Based on Geodesic for Zero-shot Cross-lingual Spoken Language Understanding
ACL 2024
·
Xuxin Cheng
Dance with Labels: Dual-Heterogeneous Label Graph Interaction for Multi-intent Spoken Language Understanding
WSDM 2024
·
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
DiffATR: Diffusion-based Generative Modeling for Audio-Text Retrieval
InterSpeech 2024
·
Yifei Xin
DBLP profile ↗
ORCID search ↗
Embracing Language Inclusivity and Diversity in CLIP through Continual Language Learning
AAAI 2024
·
Bang Yang
DBLP profile ↗
ORCID search ↗
Enhancing Dialogue State Tracking Models through LLM-backed User-Agents Simulation
ACL 2024
·
Cheng Niu
DBLP profile ↗
ORCID search ↗
Exploiting Auxiliary Caption for Video Grounding
AAAI 2024
·
Hongxiang Li
DBLP profile ↗
ORCID search ↗
Expressive Whole-Body Control for Humanoid Robots
RSS 2024
·
Xuxin Cheng
Extreme Parkour with Legged Robots
ICRA 2024
·
Xuxin Cheng
FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion Model
ACM MM 2024
·
Ziyu Yao
DBLP profile ↗
ORCID search ↗
Generating More Audios for End-to-End Spoken Language Understanding
IJCAI 2024
·
Xuxin Cheng
InMu-Net: Advancing Multi-modal Intent Detection via Information Bottleneck and Multi-sensory Processing
ACM MM 2024
·
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
KDProR: A Knowledge-Decoupling Probabilistic Framework for Video-Text Retrieval
ECCV 2024
·
Xianwei Zhuang
DBLP profile ↗
ORCID search ↗
Learning to Match Representations is Better for End-to-End Task-Oriented Dialog System
EMNLP 2024
·
Wanshi Xu
DBLP profile ↗
ORCID search ↗
MaCSC: Towards Multimodal-augmented Pre-trained Language Models via Conceptual Prototypes and Self-balancing Calibration
NAACL 2024
·
Xianwei Zhuang
DBLP profile ↗
ORCID search ↗
MoE-SLU: Towards ASR-Robust Spoken Language Understanding via Mixture-of-Experts
ACL 2024
·
Xuxin Cheng
Multivariate Cooperative Game for Image-Report Pairs: Hierarchical Semantic Alignment for Medical Report Generation
MICCAI 2024
·
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
PCAD: Towards ASR-Robust Spoken Language Understanding via Prototype Calibration and Asymmetric Decoupling
ACL 2024
·
Xianwei Zhuang
DBLP profile ↗
ORCID search ↗
PolyVoice: Language Models for Speech to Speech Translation
ICLR 2024
·
Qianqian Dong
DBLP profile ↗
ORCID search ↗
RAG-HAT: A Hallucination-Aware Tuning Pipeline for LLM in Retrieval-Augmented Generation
EMNLP 2024
·
Juntong Song
DBLP profile ↗
ORCID search ↗
Retrieval is Accurate Generation
ICLR 2024
·
Bowen Cao
DBLP profile ↗
ORCID search ↗
SaLa: Scenario-aware Label Graph Interaction for Multi-intent Spoken Language Understanding
CIKM 2024
·
Zhihong Zhu
DBLP profile ↗
ORCID search ↗
Soul-Mix: Enhancing Multimodal Machine Translation with Manifold Mixup
ACL 2024
·
Xuxin Cheng
Towards Explainable Joint Models via Information Theory for Multiple Intent Detection and Slot Filling
AAAI 2024
·
Xianwei Zhuang
DBLP profile ↗
ORCID search ↗
Towards Multi-Intent Spoken Language Understanding via Hierarchical Attention and Optimal Transport
AAAI 2024
·
Xuxin Cheng
Towards Multimodal-augmented Pre-trained Language Models via Self-balanced Expectation-Maximization Iteration
ACM MM 2024
·
Xianwei Zhuang
DBLP profile ↗
ORCID search ↗
Uncertainty-Aware Sign Language Video Retrieval with Probability Distribution Modeling
ECCV 2024
·
Xuan Wu
DBLP profile ↗
ORCID search ↗
What are the Generator Preferences for End-to-end Task-Oriented Dialog System?
EMNLP 2024
·
Wanshi Xu
DBLP profile ↗
ORCID search ↗