DE eng

Search in the Catalogues and Directories

Page: 1 2
Hits 1 – 20 of 28

1
RETRIEVING SPEAKER INFORMATION FROM PERSONALIZED ACOUSTIC MODELS FOR SPEECH RECOGNITION
In: IEEE ICASSP 2022 ; https://hal.archives-ouvertes.fr/hal-03539741 ; IEEE ICASSP 2022, 2022, Singapour, Singapore (2022)
Abstract: International audience ; The widespread of powerful personal devices capable of collecting voice of their users has opened the opportunity to build speaker adapted speech recognition system (ASR) or to participate to collaborative learning of ASR. In both cases, personalized acoustic models (AM), i.e. fine-tuned AM with specific speaker data, can be built. A question that naturally arises is whether the dissemination of personalized acoustic models can leak personal information. In this paper, we show that it is possible to retrieve the gender of the speaker, but also his identity, by just exploiting the weight matrix changes of a neural acoustic model locally adapted to this speaker. Incidentally we observe phenomena that may be useful towards explainability of deep neural networks in the context of speech processing. Gender can be identified almost surely using only the first layers and speaker verification performs well when using middle-up layers. Our experimental study on the TED-LIUM 3 dataset with HMM/TDNN models shows an accuracy of 95% for gender detection, and an Equal Error Rate of 9.07% for a speaker verification task by only exploiting the weights from personalized models that could be exchanged instead of user data.
Keyword: [INFO.INFO-CL]Computer Science [cs]/Computation and Language [cs.CL]; acoustic model; Automatic speech recognition; collaborative learning; personalized acoustic models; speaker information
URL: https://hal.archives-ouvertes.fr/hal-03539741
https://hal.archives-ouvertes.fr/hal-03539741/document
https://hal.archives-ouvertes.fr/hal-03539741/file/ICASSP_2022_SpeakerAnalysisInfoPrivacyVF.pdf
BASE
Hide details
2
Automatic Classification of Phonation Types in Spontaneous Speech: Towards a New Workflow for the Characterization of Speakers’ Voice Quality
In: Interspeech 2021 ; https://hal.archives-ouvertes.fr/hal-03334492 ; Interspeech 2021, Aug 2021, Brno, Czech Republic. pp.1015-1018, ⟨10.21437/Interspeech.2021-1765⟩ (2021)
BASE
Show details
3
Supplementary material to the paper The VoicePrivacy 2020 Challenge: Results and findings
In: https://hal.archives-ouvertes.fr/hal-03335126 ; 2021 (2021)
BASE
Show details
4
Supplementary material to the paper The VoicePrivacy 2020 Challenge: Results and findings
In: https://hal.archives-ouvertes.fr/hal-03335126 ; 2021 (2021)
BASE
Show details
5
The VoicePrivacy 2020 Challenge: Results and findings
In: https://hal.archives-ouvertes.fr/hal-03332224 ; 2021 (2021)
BASE
Show details
6
Supplementary material to the paper The VoicePrivacy 2020 Challenge: Results and findings
In: https://hal.archives-ouvertes.fr/hal-03335126 ; 2021 (2021)
BASE
Show details
7
Benchmarking and challenges in security and privacy for voice biometrics
In: SPSC 2021, 1st ISCA Symposium on Security and Privacy in Speech Communication ; https://hal.archives-ouvertes.fr/hal-03346196 ; SPSC 2021, 1st ISCA Symposium on Security and Privacy in Speech Communication, ISCA, Nov 2021, Magdeburg, Germany. ⟨10.21437/SPSC.2021-11⟩ ; https://spsc-symposium2021.de/#home (2021)
BASE
Show details
8
The VoicePrivacy 2020 Challenge: Results and findings
In: https://hal.archives-ouvertes.fr/hal-03332224 ; 2021 (2021)
BASE
Show details
9
Anonymous speaker clusters: Making distinctions between anonymised speech recordings with clustering interface
In: INTERSPEECH 2021 ; https://hal.archives-ouvertes.fr/hal-03267084 ; INTERSPEECH 2021, Aug 2021, Brno, Czech Republic (2021)
BASE
Show details
10
The VoicePrivacy 2020 Challenge Evaluation Plan
In: https://hal.archives-ouvertes.fr/hal-03623450 ; [Other] LIA - Laboratoire Informatique d'Avignon; MULTISPEECH - Speech Modeling for Facilitating Oral-Based Communication Inria Nancy - Grand Est, LORIA - NLPKD - Department of Natural Language Processing & Knowledge Discovery; Eurecom [Sophia Antipolis]; University of Edinburgh. 2020 (2020)
BASE
Show details
11
Acoustic Pairing of Original and Dubbed Voices in the Context of Video Game Localization
In: Interspeech ; https://hal.archives-ouvertes.fr/hal-01572151 ; Interspeech, Aug 2017, Stockholm, Sweden. pp.2839-2843, ⟨10.21437/Interspeech.2017-1311⟩ ; http://www.isca-speech.org/archive/Interspeech_2017/abstracts/1311.html (2017)
BASE
Show details
12
Speaker verification by inexperienced and experienced listeners vs. speaker verification system
In: IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) ; https://hal.archives-ouvertes.fr/hal-01317620 ; IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), May 2011, Prague, Czech Republic. ⟨10.1109/ICASSP.2011.5947707⟩ (2011)
BASE
Show details
13
Short Utterance-based Video Aided Speaker Recognition
In: IEEE International Workshop on Multimedia Signal Processing (MMSP) ; https://hal.archives-ouvertes.fr/hal-01927792 ; IEEE International Workshop on Multimedia Signal Processing (MMSP), Oct 2010, Cairns, France (2010)
BASE
Show details
14
Developing an acoustic-phonetic characterization of dysarthric speech in French
In: 7th International Conference on Language Resources, Technologies and Evaluation (LREC) ; https://hal.archives-ouvertes.fr/hal-00528520 ; 7th International Conference on Language Resources, Technologies and Evaluation (LREC), May 2010, Valletta, Malta. pp.2831-2838 (2010)
BASE
Show details
15
Description phonético-acoustique de la parole dysarthrique: le projet DesPho-APaDy
In: Troisièmes Journées de Phonétique Clinique ; https://halshs.archives-ouvertes.fr/halshs-00610716 ; Troisièmes Journées de Phonétique Clinique, Dec 2009, Aix en Provence, France. pp.30 (2009)
BASE
Show details
16
Back-and-Forth Methodology for Objective Voice Quality Assessment: From/to Expert Knowledge to/from Automatic Classification of Dysphonia
In: ISSN: 1687-6172 ; EISSN: 1687-6180 ; EURASIP Journal on Advances in Signal Processing ; https://hal.archives-ouvertes.fr/hal-01317140 ; EURASIP Journal on Advances in Signal Processing, SpringerOpen, 2009, 2009 (1), pp.13 - 13. ⟨10.1155/2009/982102⟩ (2009)
BASE
Show details
17
Analyse Phonétique dans le Domaine Fréquentiel pour la Classification des Voix Dysphoniques
In: Journées d'Etude sur la Parole (JEP) ; https://hal.archives-ouvertes.fr/hal-00292400 ; Journées d'Etude sur la Parole (JEP), Jun 2008, Avignon, France. pp.221-224 (2008)
BASE
Show details
18
FROM GMM TO HMM FOR EMBEDDED PASSWORD-BASED SPEAKER RECOGNITION
In: 16th European Signal Processing Conference (EUSIPCO 2008), ; https://hal.archives-ouvertes.fr/hal-01312949 ; 16th European Signal Processing Conference (EUSIPCO 2008),, Aug 2008, Lausanne, Switzerland (2008)
BASE
Show details
19
Reinforced Temporal Structure Information For Embedded Utterance-Based Speaker Recognition
In: Interspeech ; https://hal.archives-ouvertes.fr/hal-01312944 ; Interspeech, Sep 2008, brisbane, Australia (2008)
BASE
Show details
20
Characterization of the Pathological Voices (Dysphonia) in the frequency space
In: International Congress of Phonetic Sciences (ICPhS) ; https://hal.archives-ouvertes.fr/hal-00173728 ; International Congress of Phonetic Sciences (ICPhS), Aug 2007, Saarbrücken, Germany. pp.1993-1996 (2007)
BASE
Show details

Page: 1 2

Catalogues
0
0
0
0
0
0
0
Bibliographies
0
0
0
0
0
0
0
0
0
Linked Open Data catalogues
0
Online resources
0
0
0
0
Open access documents
28
0
0
0
0
© 2013 - 2024 Lin|gu|is|tik | Imprint | Privacy Policy | Datenschutzeinstellungen ändern