Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

MambAdapter: Lightweight Mamba-Based Adapters for Parameter-Efficient Transfer Learning in Speech and Audio

About

Fine-tuning Transformer-based foundation models has become the dominant strategy for domain adaptation in audio and speech processing. To reduce the computational and memory costs of this process, parameter-efficient transfer learning (PETL) methods have been widely explored. Meanwhile, Mamba, a recent state-space model, has emerged as a promising alternative to Transformers for sequence modeling. In this work, we present MambAdapter, a parameter-efficient transfer learning approach that integrates Mamba into low-rank bottleneck adapters. Our design combines parameter sharing across adapters with the injection of a lightweight Mamba module, enabling more effective modeling of audio features. We demonstrate that MambAdapter matches or outperforms strong PETL baselines on four audio classification tasks and five speech recognition languages, even when operating under reduced parameter budgets.

Salman Hussain Ali, Umberto Cappellazzo, Mirco Ravanelli• 2026

Related benchmarks

TaskDatasetResultRank
Audio ClassificationUS8K
Top-1 Accuracy82.51
30
Audio ClassificationESC
Top-1 Accuracy87.55
21
Intent PredictionFSC
Accuracy95.85
17
Speech ClassificationGSC
Accuracy94.27
11
Automatic Speech RecognitionCommonVoice 13 languages (test)--
5
Showing 5 of 5 rows

Other info

Follow for update