Szczegóły publikacji

Opis bibliograficzny

Do you read me? - flow of speech effect on speaker recognition systems / Alicja MARTINEK, Joanna Gajewska, Ewelina Bartuzi-Trokielewicz // W: Interspeech 2025 [Dokument elektroniczny] : 17–21 August 2025, Rotterdam, The Netherlands. — Wersja do Windows. — Dane tekstowe. — [France : ISCA], [2025]. — ( Interspeech : proceedings of the ... Annual Conference of the International Speech Communication Association ; ISSN  2958-1796 ). — S. 3643–3647. — Wymagania systemowe: Adobe Reader. — Tryb dostępu: https://www.isca-archive.org/interspeech_2025/martinek25_inte... [2026-01-15]. — Bibliogr. s. 3647, Abstr. — A. Martinek - dod. afiliacja: NASK National Research Institute, Poland

Autorzy (3)

Słowa kluczowe

text to speechread speechspeaker recognitionspontaneous speechspoof aware speaker verification

Dane bibliometryczne

ID BaDAP165434
Data dodania do BaDAP2026-01-15
DOI10.21437/Interspeech.2025-2629
Rok publikacji2025
Typ publikacjimateriały konferencyjne (aut.)
Otwarty dostęptak
KonferencjaInterspeech 2025
Czasopismo/seriaInterspeech

Abstract

Comparing two types of speech – read and spontaneous – poses a significant challenge for speaker verification models. This study examines the impact of these differences on the performance of advanced biometric systems. We conducted tests using two baseline speaker verification models and two spoof-aware approaches to assess their ability to handle variations between read and spontaneous speech. Additionally, we generated synthetic speech using two state-of-the-art Text-to-Speech methods, training the models either on spontaneous or read speech. The results indicate that mixing spontaneous and read speech compared with uniform type of speech yields higher error rates in biometric verification. The situation is similar in both testing scenarios, comparing genuines with impostors and genuines with synthetic speech, regardless of the type of speaker recognition model — base or spoof-aware.

Publikacje, które mogą Cię zainteresować

fragment książki
#162078Data dodania: 8.9.2025
Multi-task learning for speech emotion recognition in naturalistic conditions / Bartłomiej Zgórzyński, Juliusz Wójtowicz-Kruk, Piotr MASZTALSKI, Władysław Średniawa // W: Interspeech 2025 [Dokument elektroniczny] : 17–21 August 2025, Rotterdam, The Netherlands. — Wersja do Windows. — Dane tekstowe. — [France : ISCA], [2025]. — ( Interspeech : proceedings of the ... Annual Conference of the International Speech Communication Association ; ISSN  2958-1796 ). — S. 4678–4682. — Wymagania systemowe: Adobe Reader. — Bibliogr. s. 4682, Abstr. — P. Masztalski - dod. afiliacja: Samsung R&D Institute Poland
fragment książki
#162076Data dodania: 8.9.2025
Clustering-based hard negative sampling for supervised contrastive speaker verification / Piotr MASZTALSKI, Michał Romaniuk, Jakub Żak, Mateusz Matuszewski, Konrad KOWALCZYK // W: Interspeech 2025 [Dokument elektroniczny] : 17–21 August 2025, Rotterdam, The Netherlands. — Wersja do Windows. — Dane tekstowe. — [France : ISCA], [2025]. — ( Interspeech : proceedings of the ... Annual Conference of the International Speech Communication Association ; ISSN  2958-1796 ). — S. 3698–3702. — Wymagania systemowe: Adobe Reader. — Bibliogr. s. 3702, Abstr. — P. Masztalski - dod. afiliacja: Samsung R&D Institute Poland