Selective Attention Merging for low resource tasks: A case study of Child ASR

Shankar, Natarajan Balaji; Wang, Zilai; Eren, Eray; Alwan, Abeer

doi:10.1109/ICASSP49660.2025.10887889

Citation Details

This content will become publicly available on April 6, 2026

Selective Attention Merging for low resource tasks: A case study of Child ASR

While Speech Foundation Models (SFMs) excel in various speech tasks, their performance for low-resource tasks such as child Automatic Speech Recognition (ASR) is hampered by limited pretraining data. To address this, we explore different model merging techniques to leverage knowledge from models trained on larger, more diverse speech corpora. This paper also introduces Selective Attention (SA) Merge, a novel method that selectively merges task vectors from attention matrices to enhance SFM performance on low-resource tasks. Experiments on the MyST database show significant reductions in relative word error rate of up to 14%, outperforming existing model merging and data augmentation techniques. By combining data augmentation techniques with SA Merge, we achieve a new state-of-the-art WER of 8.69 on the MyST database for the Whisper-small model, highlighting the potential of SA Merge for improving low-resource ASR. more »

Award ID(s):: 2202585

PAR ID:: 10582855

Author(s) / Creator(s):: Shankar, Natarajan Balaji; Wang, Zilai; Eren, Eray; Alwan, Abeer

Publisher / Repository:: IEEE Signal Processing Society

Date Published:: 2025-04-06

ISBN:: 979-8-3503-6874-1

Page Range / eLocation ID:: 1 to 5

Format(s):: Medium: X

Location:: Hyderabad, India

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
This content will become publicly available on April 6, 2026
Conference Paper:
https://doi.org/10.1109/ICASSP49660.2025.10887889

More Like this