DE eng

Search in the Catalogues and Directories

Page: 1 2 3 4 5...870
Hits 1 – 20 of 17.396

1
Using Automatic Speech Recognition to Optimize Hearing-Aid Time Constants
In: ISSN: 1662-4548 ; EISSN: 1662-453X ; Frontiers in Neuroscience ; https://hal.archives-ouvertes.fr/hal-03627441 ; Frontiers in Neuroscience, Frontiers, 2022, 16 (779062), ⟨10.3389/fnins.2022.779062⟩ ; https://www.frontiersin.org/articles/10.3389/fnins.2022.779062/full (2022)
BASE
Show details
2
A fine-grained recognition of Named Entities in ELTeC collection using cascades
In: Final Action Event of COST Action Distant Reading for European Literary History ; https://hal.archives-ouvertes.fr/hal-03615219 ; Final Action Event of COST Action Distant Reading for European Literary History, Christof Schöch, Apr 2022, Krakow, Poland ; https://www.distant-reading.net/events/conference-programme/ (2022)
BASE
Show details
3
RETRIEVING SPEAKER INFORMATION FROM PERSONALIZED ACOUSTIC MODELS FOR SPEECH RECOGNITION
In: IEEE ICASSP 2022 ; https://hal.archives-ouvertes.fr/hal-03539741 ; IEEE ICASSP 2022, 2022, Singapour, Singapore (2022)
BASE
Show details
4
Emotional Speech Recognition Using Deep Neural Networks
In: ISSN: 1424-8220 ; Sensors ; https://hal.archives-ouvertes.fr/hal-03632853 ; Sensors, MDPI, 2022, 22 (4), pp.1414. ⟨10.3390/s22041414⟩ (2022)
Abstract: International audience ; The expression of emotions in human communication plays a very important role in the information that needs to be conveyed to the partner. The forms of expression of human emotions are very rich. It could be body language, facial expressions, eye contact, laughter, and tone of voice. The languages of the world’s peoples are different, but even without understanding a language in communication, people can almost understand part of the message that the other partner wants to convey with emotional expressions as mentioned. Among the forms of human emotional expression, the expression of emotions through voice is perhaps the most studied. This article presents our research on speech emotion recognition using deep neural networks such as CNN, CRNN, and GRU. We used the Interactive Emotional Dyadic Motion Capture (IEMOCAP) corpus for the study with four emotions: anger, happiness, sadness, and neutrality. The feature parameters used for recognition include the Mel spectral coefficients and other parameters related to the spectrum and the intensity of the speech signal. The data augmentation was used by changing the voice and adding white noise. The results show that the GRU model gave the highest average recognition accuracy of 97.47%. This result is superior to existing studies on speech emotion recognition with the IEMOCAP corpus.
Keyword: [INFO.INFO-AI]Computer Science [cs]/Artificial Intelligence [cs.AI]; CNN; CRNN; data augmentation; emotion; GRU; IEMOCAP; recognition; speech
URL: https://hal.archives-ouvertes.fr/hal-03632853/document
https://hal.archives-ouvertes.fr/hal-03632853
https://doi.org/10.3390/s22041414
https://hal.archives-ouvertes.fr/hal-03632853/file/sensors-22-01414-v2.pdf
BASE
Hide details
5
The Impact of Removing Head Movements on Audio-visual Speech Enhancement
In: ICASSP 2022 - IEEE International Conference on Acoustics, Speech and Signal Processing ; https://hal.inria.fr/hal-03551610 ; ICASSP 2022 - IEEE International Conference on Acoustics, Speech and Signal Processing, IEEE Signal Processing Society, May 2022, Singapore, Singapore. pp.1-5 (2022)
BASE
Show details
6
Face recognition improvements in adults and children with face recognition difficulties
Bate, S; Dalrymple, K; Bennetts, RJ. - : Oxford University Press (OUP), 2022
BASE
Show details
7
Face masks versus sunglasses: Limited effects of time and individual differences in the ability to judge facial identity and social traits
Bennetts, R; Johnson Humphrey, P; Zielinska, P. - : BioMed Central (Springer Nature) on behalf of the Psychonomic Society, 2022
BASE
Show details
8
An Overview of Indian Spoken Language Recognition from Machine Learning Perspective
In: ISSN: 2375-4699 ; EISSN: 2375-4702 ; ACM Transactions on Asian and Low-Resource Language Information Processing ; https://hal.inria.fr/hal-03616853 ; ACM Transactions on Asian and Low-Resource Language Information Processing, ACM, In press, ⟨10.1145/3523179⟩ (2022)
BASE
Show details
9
BBC-Oxford British Sign Language Dataset
In: https://hal.archives-ouvertes.fr/hal-03516444 ; 2022 (2022)
BASE
Show details
10
Fine-tuning pre-trained models for Automatic Speech Recognition: experiments on a fieldwork corpus of Japhug (Trans-Himalayan family)
In: https://halshs.archives-ouvertes.fr/halshs-03647315 ; 2022 (2022)
BASE
Show details
11
Evaluation of Speaker Anonymization on Emotional Speech ; Analyse de l'anonymisation du locuteur sur de la parole émotionnelle
In: JEP2022 - Journées d'Études sur la Parole ; https://hal.archives-ouvertes.fr/hal-03636737 ; JEP2022 - Journées d'Études sur la Parole, Jun 2022, Île de Noirmoutier, France (2022)
BASE
Show details
12
Can machines learn to see without visual databases?
In: https://hal.archives-ouvertes.fr/hal-03526569 ; 2022 (2022)
BASE
Show details
13
К вопросу о сущности основных конституционных обязанностей человека и гражданина ... : On the question of the essence of the basic constitutional duties of a person and a citizen ...
Марат Вильданович Саудаханов. - : Закон и право, 2022
BASE
Show details
14
Contextual time-continuous emotion recognition based on multimodal data ...
Fedotov, Dmitrii. - : Universität Ulm, 2022
BASE
Show details
15
Unsupervised quantification of entity consistency between photos and text in real-world news ...
Müller-Budack, Eric. - : Hannover : Institutionelles Repositorium der Leibniz Universität Hannover, 2022
BASE
Show details
16
Currencies of recognition: What rewards and recognition do Canadian distributed medical education preceptors value? ...
Johnston, Aaron. - : Open Science Framework, 2022
BASE
Show details
17
Monolinguals and Bilinguals’ Visual Recognition Memory of Socially Relevant Stimuli at 8-10 Months. ...
Freda, Kate. - : Open Science Framework, 2022
BASE
Show details
18
Dvoice : An open source dataset for Automatic Speech Recognition on African Languages and Dialects ...
BASE
Show details
19
Dvoice : An open source dataset for Automatic Speech Recognition on African Languages and Dialects ...
BASE
Show details
20
Linked Open Tafsir - Rekonstruktion der Entstehungsdynamik(en) des Korans mithilfe der Netzwerkmodellierung früher islamischer Überlieferungen ...
BASE
Show details

Page: 1 2 3 4 5...870

Catalogues
500
4
3.209
0
0
29
9
Bibliographies
9.778
0
0
0
0
0
0
2
30
Linked Open Data catalogues
0
Online resources
6
0
0
0
Open access documents
7.515
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern