Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

Towards Robust Uncertainty-Aware Speaker Modeling

About

Speaker embeddings aggregate frame-level acoustic features into compact representations for speaker recognition. Recent uncertainty-aware speaker modeling approaches further characterize the reliability of speaker embeddings by estimating their associated uncertainty. However, existing methods often suffer from inaccurate uncertainty estimation and uncertainty miscalibration under domain shifts. To address these challenges, we propose a robust uncertainty modeling framework from both estimation and adaptation perspectives. Specifically, we introduce an Inter- and Intra-Speaker-Aware Uncertainty Softmax that incorporates both inter-speaker separability and intra-speaker variability into uncertainty learning, enabling uncertainty estimates to better capture the reliability of speaker embeddings. Furthermore, we propose an Uncertainty-Calibrated Domain Adaptation (UCDA) framework to mitigate uncertainty miscalibration caused by domain mismatch. Extensive experiments on both in-domain and cross-domain benchmarks demonstrate that the proposed approach consistently improves uncertainty reliability and speaker recognition robustness.

Junjie Li, Yang Xiao, Kong Aik Lee• 2026

Related benchmarks

TaskDatasetResultRank
Speaker VerificationVoxCeleb1 (Vox1-O)
EER0.739
160
Speaker VerificationVoxCeleb1 (Vox1-H)
EER1.771
103
Speaker VerificationVoxCeleb-E
EER0.965
95
Speaker VerificationCNCeleb
EER9.237
38
Showing 4 of 4 rows

Other info

Follow for update