Back to results

Università degli studi di Trento

Emotional Prosody across languages and genders

Abstract

dc:description

This dissertation investigates emotional prosody across different languages and gen- ders, with a focus on its role in the development of Automatic Speech Emotion Recogni- tion (ASER) systems. Recognizing the limitations of existing emotional speech datasets, such as their narrow scope, limited number of speakers, and languages, this work intro- duces the Actors Challenge (AC) dataset. The AC is a dynamic, evolving, web-based interactive game designed to generate a rich speech dataset, serving as a valuable resource for studying affective prosody and other speech-related research topics. Participants not only produce emotional expressions but also evaluate the emotional performances of oth- ers, resulting in a dataset enriched with human annotations. Using one of the most advanced speech models available, the dissertation explores cross-linguistic aspects of emotional prosody, examining whether ASER systems can generalize across languages and detect cross-linguistic acoustic markers of emotion. Additionally, it investigates how gender differences play a role in expression and recognition of emotion in speech. The computational approach is complemented by a series of acoustic feature analyses, offer- ing a dual perspective that highlights the challenges and complexities of training ASER systems that accurately recognize emotions across diverse linguistic and gender contexts. The findings emphasize the complexity of emotional speech and the crucial role that di- verse, high-quality datasets play in achieving effective cross-linguistic emotion recognition in speech.

Degree

thesis:*
Grantor dc:publisher
Università degli studi di Trento
Year dc:date
2024

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Sepanta, Sia Vosh
Contributors dc:contributor
  • Zamparelli, Roberto

Subjects

dc:subject × 1

Rights

dc:rights
Statement dc:rights
  • info:eu-repo/semantics/closedAccess
  • license:Creative commons
  • license uri:http://creativecommons.org/licenses/by-nc-nd/4.0/
Language dc:language
eng

Identifiers

dc:identifier.*
OAI identifier oai:identifier
oai:iris.unitn.it:11572/438617

Chain of custody

source
Harvested from
Università degli Studi di Trento
Base URL
iris.unitn.it/oai/request
Last updated
2026-07-24
Source record
OAI-PMH GetRecord
citation

Sepanta, Sia Vosh. Emotional Prosody across languages and genders. Università degli studi di Trento, 2024. https://hdl.handle.net/11572/438617