DE eng

Search in the Catalogues and Directories

Page: 1 2 3 4 5...870
Hits 1 – 20 of 17.396

1
Using Automatic Speech Recognition to Optimize Hearing-Aid Time Constants
In: ISSN: 1662-4548 ; EISSN: 1662-453X ; Frontiers in Neuroscience ; https://hal.archives-ouvertes.fr/hal-03627441 ; Frontiers in Neuroscience, Frontiers, 2022, 16 (779062), ⟨10.3389/fnins.2022.779062⟩ ; https://www.frontiersin.org/articles/10.3389/fnins.2022.779062/full (2022)
BASE
Show details
2
A fine-grained recognition of Named Entities in ELTeC collection using cascades
In: Final Action Event of COST Action Distant Reading for European Literary History ; https://hal.archives-ouvertes.fr/hal-03615219 ; Final Action Event of COST Action Distant Reading for European Literary History, Christof Schöch, Apr 2022, Krakow, Poland ; https://www.distant-reading.net/events/conference-programme/ (2022)
BASE
Show details
3
RETRIEVING SPEAKER INFORMATION FROM PERSONALIZED ACOUSTIC MODELS FOR SPEECH RECOGNITION
In: IEEE ICASSP 2022 ; https://hal.archives-ouvertes.fr/hal-03539741 ; IEEE ICASSP 2022, 2022, Singapour, Singapore (2022)
Abstract: International audience ; The widespread of powerful personal devices capable of collecting voice of their users has opened the opportunity to build speaker adapted speech recognition system (ASR) or to participate to collaborative learning of ASR. In both cases, personalized acoustic models (AM), i.e. fine-tuned AM with specific speaker data, can be built. A question that naturally arises is whether the dissemination of personalized acoustic models can leak personal information. In this paper, we show that it is possible to retrieve the gender of the speaker, but also his identity, by just exploiting the weight matrix changes of a neural acoustic model locally adapted to this speaker. Incidentally we observe phenomena that may be useful towards explainability of deep neural networks in the context of speech processing. Gender can be identified almost surely using only the first layers and speaker verification performs well when using middle-up layers. Our experimental study on the TED-LIUM 3 dataset with HMM/TDNN models shows an accuracy of 95% for gender detection, and an Equal Error Rate of 9.07% for a speaker verification task by only exploiting the weights from personalized models that could be exchanged instead of user data.
Keyword: [INFO.INFO-CL]Computer Science [cs]/Computation and Language [cs.CL]; acoustic model; Automatic speech recognition; collaborative learning; personalized acoustic models; speaker information
URL: https://hal.archives-ouvertes.fr/hal-03539741
https://hal.archives-ouvertes.fr/hal-03539741/document
https://hal.archives-ouvertes.fr/hal-03539741/file/ICASSP_2022_SpeakerAnalysisInfoPrivacyVF.pdf
BASE
Hide details
4
Emotional Speech Recognition Using Deep Neural Networks
In: ISSN: 1424-8220 ; Sensors ; https://hal.archives-ouvertes.fr/hal-03632853 ; Sensors, MDPI, 2022, 22 (4), pp.1414. ⟨10.3390/s22041414⟩ (2022)
BASE
Show details
5
The Impact of Removing Head Movements on Audio-visual Speech Enhancement
In: ICASSP 2022 - IEEE International Conference on Acoustics, Speech and Signal Processing ; https://hal.inria.fr/hal-03551610 ; ICASSP 2022 - IEEE International Conference on Acoustics, Speech and Signal Processing, IEEE Signal Processing Society, May 2022, Singapore, Singapore. pp.1-5 (2022)
BASE
Show details
6
Face recognition improvements in adults and children with face recognition difficulties
Bate, S; Dalrymple, K; Bennetts, RJ. - : Oxford University Press (OUP), 2022
BASE
Show details
7
Face masks versus sunglasses: Limited effects of time and individual differences in the ability to judge facial identity and social traits
Bennetts, R; Johnson Humphrey, P; Zielinska, P. - : BioMed Central (Springer Nature) on behalf of the Psychonomic Society, 2022
BASE
Show details
8
An Overview of Indian Spoken Language Recognition from Machine Learning Perspective
In: ISSN: 2375-4699 ; EISSN: 2375-4702 ; ACM Transactions on Asian and Low-Resource Language Information Processing ; https://hal.inria.fr/hal-03616853 ; ACM Transactions on Asian and Low-Resource Language Information Processing, ACM, In press, ⟨10.1145/3523179⟩ (2022)
BASE
Show details
9
BBC-Oxford British Sign Language Dataset
In: https://hal.archives-ouvertes.fr/hal-03516444 ; 2022 (2022)
BASE
Show details
10
Fine-tuning pre-trained models for Automatic Speech Recognition: experiments on a fieldwork corpus of Japhug (Trans-Himalayan family)
In: https://halshs.archives-ouvertes.fr/halshs-03647315 ; 2022 (2022)
BASE
Show details
11
Evaluation of Speaker Anonymization on Emotional Speech ; Analyse de l'anonymisation du locuteur sur de la parole émotionnelle
In: JEP2022 - Journées d'Études sur la Parole ; https://hal.archives-ouvertes.fr/hal-03636737 ; JEP2022 - Journées d'Études sur la Parole, Jun 2022, Île de Noirmoutier, France (2022)
BASE
Show details
12
Can machines learn to see without visual databases?
In: https://hal.archives-ouvertes.fr/hal-03526569 ; 2022 (2022)
BASE
Show details
13
К вопросу о сущности основных конституционных обязанностей человека и гражданина ... : On the question of the essence of the basic constitutional duties of a person and a citizen ...
Марат Вильданович Саудаханов. - : Закон и право, 2022
BASE
Show details
14
Contextual time-continuous emotion recognition based on multimodal data ...
Fedotov, Dmitrii. - : Universität Ulm, 2022
BASE
Show details
15
Unsupervised quantification of entity consistency between photos and text in real-world news ...
Müller-Budack, Eric. - : Hannover : Institutionelles Repositorium der Leibniz Universität Hannover, 2022
BASE
Show details
16
Currencies of recognition: What rewards and recognition do Canadian distributed medical education preceptors value? ...
Johnston, Aaron. - : Open Science Framework, 2022
BASE
Show details
17
Monolinguals and Bilinguals’ Visual Recognition Memory of Socially Relevant Stimuli at 8-10 Months. ...
Freda, Kate. - : Open Science Framework, 2022
BASE
Show details
18
Dvoice : An open source dataset for Automatic Speech Recognition on African Languages and Dialects ...
BASE
Show details
19
Dvoice : An open source dataset for Automatic Speech Recognition on African Languages and Dialects ...
BASE
Show details
20
Linked Open Tafsir - Rekonstruktion der Entstehungsdynamik(en) des Korans mithilfe der Netzwerkmodellierung früher islamischer Überlieferungen ...
BASE
Show details

Page: 1 2 3 4 5...870

Catalogues
500
4
3.209
0
0
29
9
Bibliographies
9.778
0
0
0
0
0
0
2
30
Linked Open Data catalogues
0
Online resources
6
0
0
0
Open access documents
7.515
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern