Szczegóły publikacji

Opis bibliograficzny

Reverberant source separation using NTF with delayed subsources and spatial priors / Mieszko FRAŚ, Konrad KOWALCZYK // IEEE/ACM Transactions on Audio, Speech and Language Processing ; ISSN 2329-9290. — Tytuł poprz.: IEEE Transactions on Audio, Speech, and Language Processing ; ISSN: 1558-7916. — 2024 — vol. 32, s. 1954–1967. — Bibliogr. s. 1966, Abstr. — Publikacja dostępna online od: 2024-03-06

Autorzy (2)

Słowa kluczowe

convolutive nonnegative matrix factorizationsource separationroom reverberationarray signal processing

Dane bibliometryczne

ID BaDAP154436
Data dodania do BaDAP2024-07-15
Tekst źródłowyURL
DOI10.1109/TASLP.2024.3374065
Rok publikacji2024
Typ publikacjiartykuł w czasopiśmie
Otwarty dostęptak
Czasopismo/seriaIEEE/ACM Transactions on Audio, Speech and Language Processing

Abstract

Speech signals recorded by distant microphones are often contaminated with room reverberation and signals of interfering speakers. This article addresses the problem of joint source separation and dereverberation using multichannel nonnegative tensor factorization (NTF) in which late reverberant components are modeled using the so-called delayed subsources. The article formulates two distinct signal models of the time-frequency spectrum of the multichannel microphone mixture, in which reverberation is modeled either independently for each source using delayed source variances or jointly using delayed microphone signals. In addition, it defines computationally efficient variants of these two methods with a simplified spatial model in which spatial properties of the late reverberant components are estimated jointly for all delays. For each of the four distinct algorithms, the article first formulates a maximum a posteriori (MaP) estimator based on the NTF model with the localization prior over the mixing matrix that is suitable for the estimation of the early reverberation (primarily the direct-path) signals in a reverberant environment. Next it derives update equations for the four resulting expectation-maximization algorithms, which are thoroughly evaluated and shown to outperform similar state-of-the-art approaches. The results of experimental evaluations, performed using real and simulated data, for determined, over-determined and under-determined scenarios, indicate superior performance of the proposed processing over state-of-the-art in terms of standard source separation and dereverberation metrics.

Publikacje, które mogą Cię zainteresować

artykuł
#154433Data dodania: 15.7.2024
On ambisonic source separation with spatially informed non-negative tensor factorization / Mateusz GUZIK, Konrad KOWALCZYK // IEEE/ACM Transactions on Audio, Speech and Language Processing ; ISSN 2329-9290. — Tytuł poprz.: IEEE Transactions on Audio, Speech, and Language Processing ; ISSN: 1558-7916. — 2024 — vol. 32, s. 3238–3255. — Bibliogr. s. 3254–3255, Abstr. — Publikacja dostępna online od: 2024-05-10
fragment książki
#146975Data dodania: 2.6.2023
Convolutive NTF for ambisonic source separation under reverberant conditions / Mateusz GUZIK, Konrad KOWALCZYK // W: ICASSP 2023 [Dokument elektroniczny] : 2023 IEEE International Conference on Acoustics, Speech and Signal Processing : 4–10 June, Rhodes Island, Greece : conference proceedings. — Wersja do Windows. — Dane tekstowe. — Piscataway : IEEE, cop. 2023. — e-ISBN: 978-1-7281-6327-7. — S. [1–5]. — Wymagania systemowe: Adobe Reader. — Bibliogr. s. 5, Abstr. — Publikacja dostępna online od: 2023-05-05