Szczegóły publikacji

Opis bibliograficzny

Causal signal-based DCCRN with overlapped-frame prediction for online speech enhancement / Julitta BARTOLEWSKA, Stanisław KACPRZAK, Konrad KOWALCZYK // W: INTERSPEECH 2023 [Dokument elektroniczny] : Dublin, Ireland, 20-24 August 2023 / [eds.] Naomi Harte, Julie Carson-Berndsen, Gareth Jones. — Wersja do Windows. — Dane tekstowe. — [France : ISCA], [2023]. — S. 4039–4043. — Wymagania systemowe: Adobe Reader. — Tryb dostępu: https://www.isca-speech.org/archive/pdfs/interspeech_2023/bar... [2023-09-19]. — Bibliogr. s. 4043, Abstr.

Autorzy (3)

Słowa kluczowe

online processingdeep neural networknoise suppressionspeech enhancement

Dane bibliometryczne

ID BaDAP148720
Data dodania do BaDAP2023-09-20
DOI10.21437/Interspeech.2023-2177
Rok publikacji2023
Typ publikacjimateriały konferencyjne (aut.)
Otwarty dostęptak
KonferencjaInterspeech 2023

Abstract

The aim of speech enhancement is to improve speech signal quality and intelligibility from a noisy microphone signal. In many applications, it is crucial to enable processing with small computational complexity and minimal requirements regarding access to future signal samples (look-ahead). This paper presents signal-based causal DCCRN that improves online single-channel speech enhancement by reducing the required look-ahead and the number of network parameters. The proposed modifications include complex filtering of the signal, application of overlapped-frame prediction, causal convolutions and deconvolutions, and modification of the loss function. Results of performed experiments indicate that the proposed model with overlapped signal prediction and additional adjustments, achieves similar or better performance than the original DCCRN in terms of various speech enhancement metrics, while it reduces the latency and network parameter number by around 30%.

Publikacje, które mogą Cię zainteresować

fragment książki
#148719Data dodania: 20.9.2023
Joint blind source separation and dereverberation for automatic speech recognition using delayed-subsource MNMF with localization prior / Mieszko FRAŚ, Marcin WITKOWSKI, Konrad KOWALCZYK // W: INTERSPEECH 2023 [Dokument elektroniczny] : Dublin, Ireland, 20-24 August 2023 / [eds.] Naomi Harte, Julie Carson-Berndsen, Gareth Jones. — Wersja do Windows. — Dane tekstowe. — [France : ISCA], [2023]. — S. 3734–3738. — Wymagania systemowe: Adobe Reader. — Tryb dostępu: https://www.isca-speech.org/archive/pdfs/interspeech_2023/fra... [2023-09-19]. — Bibliogr. s. 3738, Abstr.
fragment książki
#162187Data dodania: 10.9.2025
Efficient low-latency speech enhancement with mobile audio streaming networks / Michal Romaniuk, Piotr Masztalski, Karol Piaskowski, Mateusz Matuszewski // W: INTERSPEECH 2020 [Dokument elektroniczny] : October 25–29, Shanghai, China. — Wersja do Windows. — Dane tekstowe. — [China] : ISCA, cop. 2020. — ( Interspeech : proceedings of the ... Annual Conference of the International Speech Communication Association ; ISSN  2958-1796 ). — S. 3296-3300. — Wymagania systemowe: Adobe Reader. — Tryb dostępu: https://www.isca-archive.org/interspeech_2020/romaniuk20_inte... [2025-09-09]. — Bibliogr. s. 3299-3300, Abstr.