1 |
A New Amharic Speech Emotion Dataset and Classification Benchmark ...
|
|
|
|
Abstract:
In this paper we present the Amharic Speech Emotion Dataset (ASED), which covers four dialects (Gojjam, Wollo, Shewa and Gonder) and five different emotions (neutral, fearful, happy, sad and angry). We believe it is the first Speech Emotion Recognition (SER) dataset for the Amharic language. 65 volunteer participants, all native speakers, recorded 2,474 sound samples, two to four seconds in length. Eight judges assigned emotions to the samples with high agreement level (Fleiss kappa = 0.8). The resulting dataset is freely available for download. Next, we developed a four-layer variant of the well-known VGG model which we call VGGb. Three experiments were then carried out using VGGb for SER, using ASED. First, we investigated whether Mel-spectrogram features or Mel-frequency Cepstral coefficient (MFCC) features work best for Amharic. This was done by training two VGGb SER models on ASED, one using Mel-spectrograms and the other using MFCC. Four forms of training were tried, standard cross-validation, and ... : 16 pages, 12 tables, 6 figures ...
|
|
Keyword:
Audio and Speech Processing eess.AS; Computation and Language cs.CL; FOS Computer and information sciences; FOS Electrical engineering, electronic engineering, information engineering; I.2.7; Sound cs.SD
|
|
URL: https://dx.doi.org/10.48550/arxiv.2201.02710 https://arxiv.org/abs/2201.02710
|
|
BASE
|
|
Hide details
|
|
2 |
A Deep CNN Architecture with Novel Pooling Layer Applied to Two Sudanese Arabic Sentiment Datasets ...
|
|
|
|
BASE
|
|
Show details
|
|
3 |
A Novel Interval-Valued q-Rung Dual Hesitant Linguistic Multi-Attribute Decision-Making Method Based on Linguistic Scale Functions and Power Hamy Mean
|
|
|
|
In: Entropy; Volume 24; Issue 2; Pages: 166 (2022)
|
|
BASE
|
|
Show details
|
|
4 |
MKPM: Multi keyword-pair matching for natural language sentences
|
|
|
|
BASE
|
|
Show details
|
|
7 |
Additional file 1: of Change patterns of oncomelanid snail burden in areas within the Yangtze River drainage after the three gorges dam operated ...
|
|
|
|
BASE
|
|
Show details
|
|
8 |
Additional file 1: of Imported malaria cases in former endemic and non-malaria endemic areas in China: are there differences in case profile and time to response? ...
|
|
|
|
BASE
|
|
Show details
|
|
9 |
Additional file 1: of Change patterns of oncomelanid snail burden in areas within the Yangtze River drainage after the three gorges dam operated ...
|
|
|
|
BASE
|
|
Show details
|
|
10 |
Additional file 1: of Imported malaria cases in former endemic and non-malaria endemic areas in China: are there differences in case profile and time to response? ...
|
|
|
|
BASE
|
|
Show details
|
|
11 |
No Morphological Markers, No Problem: ERP Study Reveals Semantic Contribution to Distinct Neural Substrates Between Noun and Verb Processing in Online Sentence Comprehension
|
|
|
|
BASE
|
|
Show details
|
|
12 |
Additional file 1: of A case report of spontaneous abortion caused by Brucella melitensis biovar 3 ...
|
|
|
|
BASE
|
|
Show details
|
|
13 |
Additional file 1: of A case report of spontaneous abortion caused by Brucella melitensis biovar 3 ...
|
|
|
|
BASE
|
|
Show details
|
|
15 |
A Study on the Integration Risk Management for the Insurance Enterprises
|
|
|
|
In: Cross-Cultural Communication; Vol 4, No 3 (2008): Cross-Cultural Communication; 44-50 ; 1923-6700 ; 1712-8358 (2010)
|
|
BASE
|
|
Show details
|
|
|
|