Szczegóły publikacji

Opis bibliograficzny

Multi-task learning for speech emotion recognition in naturalistic conditions / Bartłomiej Zgórzyński, Juliusz Wójtowicz-Kruk, Piotr MASZTALSKI, Władysław Średniawa // W: Interspeech 2025 [Dokument elektroniczny] : 17–21 August 2025, Rotterdam, The Netherlands. — Wersja do Windows. — Dane tekstowe. — [France : ISCA], [2025]. — ( Interspeech : proceedings of the ... Annual Conference of the International Speech Communication Association ; ISSN  2958-1796 ). — S. 4678–4682. — Wymagania systemowe: Adobe Reader. — Bibliogr. s. 4682, Abstr. — P. Masztalski - dod. afiliacja: Samsung R&D Institute Poland

Autorzy (4)

Słowa kluczowe

multi-modal learningspeech emotion recognitionmulti-task learning

Dane bibliometryczne

ID BaDAP162078
Data dodania do BaDAP2025-09-08
Tekst źródłowyURL
DOI10.21437/Interspeech.2025-2033
Rok publikacji2025
Typ publikacjimateriały konferencyjne (aut.)
Otwarty dostęptak
KonferencjaInterspeech 2025
Czasopismo/seriaInterspeech

Abstract

This work introduces a multi-encoder joint classification and regression training framework for speech emotion recognition. We present our solution for the Interspeech 2025 Speech Emotion Recognition in Naturalistic Conditions Challenge, leveraging a multi-modal, multi-encoder architecture with a fusion module. Our results demonstrate the effectiveness of the multi-task approach for both classification and regression tasks, achieving a top 10 spot in categorical emotion classification and 2nd place in emotional attribute prediction among competing teams. Furthermore, an ablation study shows that employing multi-task learning outperforms separate task-specific training. These findings highlight the potential of multi-task, multi-encoder systems for speech emotion recognition.

Publikacje, które mogą Cię zainteresować

fragment książki
#165876Data dodania: 6.3.2026
Leveraging text and speech processing for suicide risk classification in Chinese adolescents / Justyna Krzywdziak, Bartłomiej Eljasiak, Joanna Stępień, Michał Świątek, Agnieszka Pruszek // W: Interspeech 2025 [Dokument elektroniczny] : 17–21 August 2025, Rotterdam, The Netherlands. — Wersja do Windows. — Dane tekstowe. — [France : ISCA], [2025]. — ( Interspeech : proceedings of the ... Annual Conference of the International Speech Communication Association ; ISSN  2958-1796 ). — S. 394–398. — Wymagania systemowe: Adobe Reader. — Bibliogr. s. 398, Abstr. — J. Krzywdziak - afiliacja: Samsung R&D Institute Poland
fragment książki
#165434Data dodania: 15.1.2026
Do you read me? - flow of speech effect on speaker recognition systems / Alicja MARTINEK, Joanna Gajewska, Ewelina Bartuzi-Trokielewicz // W: Interspeech 2025 [Dokument elektroniczny] : 17–21 August 2025, Rotterdam, The Netherlands. — Wersja do Windows. — Dane tekstowe. — [France : ISCA], [2025]. — ( Interspeech : proceedings of the ... Annual Conference of the International Speech Communication Association ; ISSN  2958-1796 ). — S. 3643–3647. — Wymagania systemowe: Adobe Reader. — Tryb dostępu: https://www.isca-archive.org/interspeech_2025/martinek25_inte... [2026-01-15]. — Bibliogr. s. 3647, Abstr. — A. Martinek - dod. afiliacja: NASK National Research Institute, Poland