Low-Delay Speech Enhancement Using Perceptually Motivated Target and Loss

Speech enhancement approaches based on deep neural network have outperformed the traditional signal processing methods. This paper presents a low-delay speech enhancement method that employs a new perceptually motivated training target and loss function. The proposed approach can achieve similar speech enhancement performance compared to the state-of-the-art approaches, but with significantly less latency and computational complexities. Judged by the MOS tests conducted by the INTERSPEECH 2021 Deep Noise Suppression Challenge organizer, the proposed method is ranked the 2 nd place for Background Noise MOS, and the 6 th place for overall MOS.

Paper

Full text

PDF

Low-Delay Speech Enhancement Using Perceptually Motivated Target and Loss

Semantic Scholar · Computer Science · 2021

Abstract

Speech enhancement approaches based on deep neural network have outperformed the traditional signal processing methods. This paper presents a low-delay speech enhancement method that employs a new perceptually motivated training target and loss function. The proposed approach can achieve similar speech enhancement performance compared to the state-of-the-art approaches, but with significantly less latency and computational complexities. Judged by the MOS tests conducted by the INTERSPEECH 2021 Deep Noise Suppression Challenge organizer, the proposed method is ranked the 2 nd place for Background Noise MOS, and the 6 th place for overall MOS.

Similar papers

© 2026 NYSGPT2525 LLC