Progressive Subband Modeling for Artifacts-free Speech Super-resolution

Citations

WEB OF SCIENCE

0
Citations

SCOPUS

0

초록

In this paper, we consider new reconstruction loss together with a subband objective in the form of auxiliary loss function for artifacts-free speech super-resolution. Unlike prior work which mainly consider full band of frequency region for speech super-resolution, the proposed method alleviates distortion generated during deep learning training via subband modeling. To further minimize spectral artifacts, we also apply progressive curriculum learning for superior performance. Our experimental results demonstrate that the proposed method outperforms the evaluated baselines on the both TIMIT and VCTK dataset by increase in both intelligibility and perceptual score. Furthermore, the visual representation of spectrograms comparison verify that our proposed method clearly restoring speech with fewer artifacts. Audio samples and the implementations are available online.

키워드

bandwidth extensioncurriculum learningspectral artifactsspeech super-resolutionBandwidthComputer visionGearsIntelligent systemsSpeech communication
제목
Progressive Subband Modeling for Artifacts-free Speech Super-resolution
저자
Kim, DonghyunChang, Joon-Hyuk
DOI
10.1109/ICASSP49660.2025.10889911
발행일
2025-03
유형
Conference paper
저널명
ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings
페이지
1 ~ 5