1 |
Some issues affecting the transcription of hungarian broadcast audio
|
|
|
|
In: Annual Conference of the International Speech Communication Association ; https://hal.archives-ouvertes.fr/hal-01843430 ; Annual Conference of the International Speech Communication Association , Aug 2013, Lyon, France (2013)
|
|
BASE
|
|
Show details
|
|
2 |
Acoustic unit discovery and pronunciation generation from a grapheme-based lexicon
|
|
|
|
In: IEEE Automatic Speech Recognition and Understanding Workshop ; https://hal.archives-ouvertes.fr/hal-01843433 ; IEEE Automatic Speech Recognition and Understanding Workshop, Dec 2013, Olomouc, Czech Republic (2013)
|
|
BASE
|
|
Show details
|
|
3 |
Discriminative training of a phoneme confusion model for a dynamic lexicon in ASR
|
|
|
|
In: Interspeech 2013 ; Annual Conference of the International Speech Communication Association ; https://hal.archives-ouvertes.fr/hal-01843427 ; Annual Conference of the International Speech Communication Association, Jan 2013, Lyon, France (2013)
|
|
Abstract:
International audience ; To enhance the recognition lexicon, it is important to be able to add pronunciation variants while keeping the confusability introduced by the extra phonemic variation low. However, this confusability is not easily correlated with the ASR performance, as it is an inherent phenomenon of speech. This paper proposes a method to construct a multiple pronunciation lexicon with a high discriminability. To do so, a phoneme confusion model is used to expand the phonemic search space of pronunciation variants during ASR decoding and a discriminative framework is adopted for the training of the weights of the phoneme confusions. For the parameter estimation, two training algorithms are implemented, the perceptron and the CRF model, using finite state transducers. Experiments on English data were conducted using a large state-of-the-art ASR system of continuous speech.
|
|
Keyword:
[INFO.INFO-CL]Computer Science [cs]/Computation and Language [cs.CL]; [INFO]Computer Science [cs]; discriminative training; dynamic recognition lexicon; FST-based ASR decoding; phoneme confusion model
|
|
URL: https://hal.archives-ouvertes.fr/hal-01843427 https://hal.archives-ouvertes.fr/hal-01843427/document https://hal.archives-ouvertes.fr/hal-01843427/file/karanasou13_interspeech.pdf
|
|
BASE
|
|
Hide details
|
|
4 |
Recent evolution of non-standard consonantal variants in French broadcast news
|
|
|
|
In: Interspeech ; https://halshs.archives-ouvertes.fr/halshs-00856290 ; Interspeech, Aug 2013, Lyon, France. pp.412-416 (2013)
|
|
BASE
|
|
Show details
|
|
5 |
Recent Evolution of Non Standard Consonantal Variants in French Broadcast News
|
|
|
|
In: Annual Conference of the International Speech Communication Association ; https://hal.archives-ouvertes.fr/hal-01843431 ; Annual Conference of the International Speech Communication Association , International Speech Communication Association, F. Bimbot, C. Cerisara, C. Fougeron, G. Gravier, L. Lamel, F. Pellegrino, P. Perrier, Jan 2013, Lyon, France (2013)
|
|
BASE
|
|
Show details
|
|
6 |
Unsupervised Acoustic Model Training with Limited Linguistic Resources
|
|
|
|
In: IEEE Automatic Speech Recognition and Understanding Workshop ; https://hal.archives-ouvertes.fr/hal-01843476 ; IEEE Automatic Speech Recognition and Understanding Workshop, Jan 2013, Olomouc, Czech Republic (2013)
|
|
BASE
|
|
Show details
|
|
7 |
What we can learn from ASR errors about low-resourced languages: a case- study of Luxembourgish and Austrian
|
|
|
|
In: Errors by Humans and Machines in Multimedia, Multimodal, Multilingual Data Processing ; https://hal.archives-ouvertes.fr/hal-01843440 ; Errors by Humans and Machines in Multimedia, Multimodal, Multilingual Data Processing, Jan 2013, Ermenonville, France (2013)
|
|
BASE
|
|
Show details
|
|
8 |
Embosi: automatic alignment with segments and words and phonological mining
|
|
|
|
In: International Conference on Bantu Languages ; https://hal.archives-ouvertes.fr/hal-01843438 ; International Conference on Bantu Languages, Jan 2013, Paris, France (2013)
|
|
BASE
|
|
Show details
|
|
9 |
What we can learn from asr errors about low-resourced languages: a case-study of luxembourgish and austrian
|
|
|
|
In: Errors by Humans and Machines in Multimedia, Multimodal, Multilingual Data Processing (ERRARE 2013) ; https://halshs.archives-ouvertes.fr/halshs-01424902 ; Errors by Humans and Machines in Multimedia, Multimodal, Multilingual Data Processing (ERRARE 2013), Nov 2013, Ermenonville, France (2013)
|
|
BASE
|
|
Show details
|
|
10 |
Embosi : automatic alignment with segments and words and phonological mining
|
|
|
|
In: International Conference on Bantu Languages (BANTU 2013) ; https://halshs.archives-ouvertes.fr/halshs-01424894 ; International Conference on Bantu Languages (BANTU 2013), Jun 2013, Paris France (2013)
|
|
BASE
|
|
Show details
|
|
11 |
Human annotation of asr error regions: Is ”gravity” a sharable concept for human annotators?
|
|
|
|
In: Errors by Humans and Machines in Multimedia, Multimodal, Multilingual Data Processing (ERRARE 2013) ; https://halshs.archives-ouvertes.fr/halshs-01424915 ; Errors by Humans and Machines in Multimedia, Multimodal, Multilingual Data Processing (ERRARE 2013), Nov 2013, Ermenonville, France (2013)
|
|
BASE
|
|
Show details
|
|
12 |
Systèmes de transcription comme instruments
|
|
|
|
In: Méthodes et outils pour l'analyse phonétique des grands corpus oraux ; https://hal.archives-ouvertes.fr/hal-01135113 ; Nguyen Noël; Adda-Decker Martine. Méthodes et outils pour l'analyse phonétique des grands corpus oraux, Hermes Science Publications, pp.159-202, 2013, Cognition et Traitement de l'Information, 978-2746245303 (2013)
|
|
BASE
|
|
Show details
|
|
13 |
Proceedings of the 14th Annual Conference of the International Speech Communication Association (Interspeech), 25-29 August 2013, Lyon (France)
|
|
|
|
In: https://hal.archives-ouvertes.fr/hal-00931864 ; France. International Speech Communication Association (ISCA), over 3500 p., 2013 (2013)
|
|
BASE
|
|
Show details
|
|
|
|