P
PaperPicks
Conferences
← All conferences
Editions
2026
2025
2024
InterSpeech
2025
A
Conference of the International Speech Communication Association
Official site ↗
1,181
accepted papers
1,099
authors
August 17-22, 2025
dates
Netherlands
location
1,181
/ 1,181 papers
All
1,181
Main Track
1,180
Editorship
1
1
"Alexa, can you forget me?" Machine Unlearning Benchmark in Spoken Language Understanding
Alkis Koudounas
DBLP profile ↗
ORCID search ↗
2
"Dyadosyncrasy", Idiosyncrasy and Demographic Factors in Turn-Taking
Julio Cesar Cavalcanti
DBLP profile ↗
ORCID search ↗
3
"KAN you hear me?" Exploring Kolmogorov-Arnold Networks for Spoken Language Understanding
Alkis Koudounas
DBLP profile ↗
ORCID search ↗
4
26th Annual Conference of the International Speech Communication Association, Interspeech 2025, Rotterdam, The Netherlands, 17-21 August 2025
Odette Scharenborg
DBLP profile ↗
ORCID search ↗
5
2D Immersed Boundary Method in Vocal Tract Acoustics: An Eulerian-Lagrangian Model for Simulation of Diphthongs
Rongshuai Wu
DBLP profile ↗
ORCID search ↗
6
75-Speaker Annot-16: A benchmark dataset for speech articulatory rt-MRI annotation with articulator contours and phonetic alignment
Xuan Shi
DBLP profile ↗
ORCID search ↗
7
A Bayesian Approach to L2 Fluency Ratings by Native and Nonnative Listeners
Kakeru Yazawa
DBLP profile ↗
ORCID search ↗
8
A Cascaded Multimodal Framework for Automatic Social Communication Severity Assessment in Children with Autism Spectrum Disorder
Jihyun Mun
DBLP profile ↗
ORCID search ↗
9
A Chinese Heart Failure Status Speech Database with Universal and Personalised Classification
Yue Pan
DBLP profile ↗
ORCID search ↗
10
A Comparative Study on Proactive and Passive Detection of Deepfake Speech
Chia-Hua Wu
DBLP profile ↗
ORCID search ↗
11
A Comprehensive Real-World Assessment of Audio Watermarking Algorithms: Will They Survive Neural Codecs?
Yigitcan Özer
DBLP profile ↗
ORCID search ↗
12
A Cookbook for Community-driven Data Collection of Impaired Speech in Low-Resource Languages
Sumaya Ahmed Salihs
DBLP profile ↗
ORCID search ↗
13
A Copula-Based Generative Score-Level Fusion Model for Speaker Verification
Sandro Cumani
DBLP profile ↗
ORCID search ↗
14
A Data-Driven Diffusion-based Approach for Audio Deepfake Explanations
Petr Grinberg
DBLP profile ↗
ORCID search ↗
15
A Dataset for Automatic Assessment of TTS Quality in Spanish
Alejandro Sosa Welford
DBLP profile ↗
ORCID search ↗
16
A Deformable Convolution GAN Approach for Speech Dereverberation in Cochlear Implant Users
Hsin-Tien Chiang
DBLP profile ↗
ORCID search ↗
17
A Domain Robust Pre-Training Method with Local Prototypes for Speaker Verification
Qing Gu
DBLP profile ↗
ORCID search ↗
18
A Gradient Effect of Hand Beat Timing on Spoken Word Recognition
Chengjia Ye
DBLP profile ↗
ORCID search ↗
19
A Hybrid Approach to Combining Role Diarization with ASR for Professional Conversations
Bongjun Kim
DBLP profile ↗
ORCID search ↗
20
A Joint Network for Singing Melody Extraction from Polyphonic Music with Attention Aggregation and Self-Consistency Training
Jiabo Jing
DBLP profile ↗
ORCID search ↗
21
A Lightweight Hybrid Dual Channel Speech Enhancement System under Low-SNR Conditions
Zheng Wang
DBLP profile ↗
ORCID search ↗
22
A Multi-Dialectal Dataset for German Dialect ASR and Dialect-to-Standard Speech Translation
Verena Blaschke
DBLP profile ↗
ORCID search ↗
23
A Multi-Stream Framework Utilizing 3D Human Reconstruction for Cued Speech Recognition
Katerina Papadimitriou
DBLP profile ↗
ORCID search ↗
24
A Multimodal Chinese Dataset for Cross-lingual Sarcasm Detection
Xiyuan Gao
DBLP profile ↗
ORCID search ↗
25
A Naturally Elicited Multimodal Stress Database and Speech Breathing Based Stress Detection
Karumannil Mohamed Ismail Yasar Arafath
DBLP profile ↗
ORCID search ↗
26
A Neural Codec Approach for Noise-Robust Bandwidth Expansion
Xi Liu
DBLP profile ↗
ORCID search ↗
27
A Novel Deep Learning Framework for Efficient Multichannel Acoustic Feedback Control
Yuan-Kuei Wu
DBLP profile ↗
ORCID search ↗
28
A Perception-Based L2 Speech Intelligibility Indicator: Leveraging a Rater's Shadowing and Sequence-to-sequence Voice Conversion
Haopeng Geng
DBLP profile ↗
ORCID search ↗
29
A Practitioner's Guide to Building ASR Models for Low-Resource Languages: A Case Study on Scottish Gaelic
Ondrej Klejch
DBLP profile ↗
ORCID search ↗
30
A real-time MRI study on asymmetry in velum dynamics during VCV production with nasal sounds
Chetan Sharma
DBLP profile ↗
ORCID search ↗
31
A Robust Hybrid ACC-PM Approach for Personal Sound Zones
Yaqi Zhu
DBLP profile ↗
ORCID search ↗
32
A Self-Training Approach for Whisper to Enhance Long Dysarthric Speech Recognition
Shiyao Wang
DBLP profile ↗
ORCID search ↗
33
A Semantic Information-based Hierarchical Speech Enhancement Method Using Factorized Codec and Diffusion Model
Yang Xiang
DBLP profile ↗
ORCID search ↗
34
A semi-automatic pipeline for transcribing and segmenting child speech
Polychronia Christodoulidou
DBLP profile ↗
ORCID search ↗
35
A Siamese Network-Based Framework for Voice Mimicry Proficiency Assessment Using X-Vector Embeddings
Bhasi K. C.
DBLP profile ↗
ORCID search ↗
36
A Silent Speech Decoding System from EEG and EMG with Heterogenous Electrode Configurations
Masakazu Inoue
DBLP profile ↗
ORCID search ↗
37
A simple method for predicting Clinical Scores in Huntington's Disease by leveraging ASR's uncertainty on spontaneous speech
Hadrien Titeux
DBLP profile ↗
ORCID search ↗
38
A Simple-Yet-Effective Data Augmentation Method for Speaker Identification in Novels
Wenjie Zhong
DBLP profile ↗
ORCID search ↗
39
A Study of Real-world Audio-Visual Corpus Design and Production: A Perspective from MISP Challenges
Hang Chen
DBLP profile ↗
ORCID search ↗
40
A Study of Speech Embedding Similarities Between Australian Aboriginal and High-Resource Languages
Eliathamby Ambikairajah
DBLP profile ↗
ORCID search ↗
41
A Study on Speech Assessment with Visual Cues
Shafique Ahmed
DBLP profile ↗
ORCID search ↗
42
A Study on The Impact of Foundation Models on Automatic Depression Detection from Speech Signals
Bubai Maji
DBLP profile ↗
ORCID search ↗
43
A Three-Stage Beamforming with Harmonic Guidance for Multi-Channel Speech Enhancement
Nurali Alip
DBLP profile ↗
ORCID search ↗
44
A Two-Stage Hierarchical Deep Filtering Framework for Real-Time Speech Enhancement
Shenghui Lu
DBLP profile ↗
ORCID search ↗
45
A Watermark for Auto-Regressive Speech Generation Models
Yihan Wu
DBLP profile ↗
ORCID search ↗
46
A-SMiLE: Affective Sparse Mixture-of-Experts Adapter with Multi-Task Learning for Spoken Dialogue Models
Yi-Wen Chao
DBLP profile ↗
ORCID search ↗
47
AA-SLLM: An Acoustically Augmented Speech Large Language Model for Speech Emotion Recognition
Jialong Mai
DBLP profile ↗
ORCID search ↗
48
ABHINAYA - A System for Speech Emotion Recognition In Naturalistic Conditions Challenge
Soumya Dutta
DBLP profile ↗
ORCID search ↗
49
AC/DC: LLM-based Audio Comprehension via Dialogue Continuation
Yusuke Fujita
DBLP profile ↗
ORCID search ↗
50
Accelerating Autoregressive Speech Synthesis Inference With Speech Speculative Decoding
Zijian Lin
DBLP profile ↗
ORCID search ↗
Show 100 more
(1,131 left)