Print Email Facebook Twitter A Cross-Corpus Speech-Based Analysis of Escalating Negative Interactions Title A Cross-Corpus Speech-Based Analysis of Escalating Negative Interactions Author Lefter, I. (TU Delft System Engineering) Baird, Alice (Universität Augsburg) Stappen, Lukas (Universität Augsburg) Schuller, Björn W. (Imperial College London; Universität Augsburg) Date 2022 Abstract The monitoring of an escalating negative interaction has several benefits, particularly in security, (mental) health, and group management. The speech signal is particularly suited to this, as aspects of escalation, including emotional arousal, are proven to easily be captured by the audio signal. A challenge of applying trained systems in real-life applications is their strong dependence on the training material and limited generalization abilities. For this reason, in this contribution, we perform an extensive analysis of three corpora in the Dutch language. All three corpora are high in escalation behavior content and are annotated on alternative dimensions related to escalation. A process of label mapping resulted in two possible ground truth estimations for the three datasets as low, medium, and high escalation levels. To observe class behavior and inter-corpus differences more closely, we perform acoustic analysis of the audio samples, finding that derived labels perform similarly across each corpus, with escalation interaction increasing in pitch (F0) and intensity (dB). We explore the suitability of different speech features, data augmentation, merging corpora for training, and testing on actor and non-actor speech through our experiments. We find that the extent to which merging corpora is successful depends greatly on the similarities between label definitions before label mapping. Finally, we see that the escalation recognition task can be performed in a cross-corpus setup with hand-crafted speech features, obtaining up to 63.8% unweighted average recall (UAR) at best for a cross-corpus analysis, an increase from the inter-corpus results of 59.4% UAR. Subject affective computingnegative interactionscross-corpora analysisconflict escalationspeech paralinguisticsemotion recognition To reference this document use: http://resolver.tudelft.nl/uuid:9a85607d-9797-4b1d-af9b-254e6bdad0d2 DOI https://doi.org/10.3389/fcomp.2022.749804 ISSN 2624-9898 Source Frontiers in Computer Science, 4 Part of collection Institutional Repository Document type journal article Rights © 2022 I. Lefter, Alice Baird, Lukas Stappen, Björn W. Schuller Files PDF Fcomp_04_749804.pdf 1.84 MB Close viewer /islandora/object/uuid:9a85607d-9797-4b1d-af9b-254e6bdad0d2/datastream/OBJ/view