상세 보기
Language Model Personalization for Speech Recognition: A Clustered Federated Learning Approach with Adaptive Weight Average
- Lee, Chae-Won;
- Lee, Jae-Hong;
- Chang, Joon-Hyuk
WEB OF SCIENCE
0SCOPUS
2초록
In the rapidly evolving field of automatic speech recognition (ASR), the push towards personalization has become a paramount concern. Text-only personalization, while advantageous for data collection and adaptable to text variations, can suffer from overfitting when using personal data and requires extensive data to mitigate this issue. Federated learning (FL) emerges as a solution, facilitating learning from diverse client models while preserving privacy. However, FL addresses the challenges posed by non independent and identically distributed (non-i.i.d) data, potentially leading to poor performance. We propose two approaches for language model personalization in ASR to address these issues. First, adaptive weighted average addresses the limitations of uniform weight average in the existing FL method by combining local language models into a global model. Second, clustered federated learning, based solely on model parameters, improves model stability without relying on information from the local domain. Both strategies aim to enhance personalization and reduce performance degradation, particularly in non-i.i.d scenarios within the FL.
키워드
- 제목
- Language Model Personalization for Speech Recognition: A Clustered Federated Learning Approach with Adaptive Weight Average
- 저자
- Lee, Chae-Won; Lee, Jae-Hong; Chang, Joon-Hyuk
- 발행일
- 2024-07
- 유형
- Article
- 권
- 31
- 페이지
- 2710 ~ 2714