CRIM'S FRENCH SPEECH TRANSCRIPTION SYSTEM FOR ETAPE 2011

被引:0
|
作者
Gupta, Vishwa [1 ]
Boulianne, Gilles [1 ]
Osterrath, Frederic [1 ]
Ouellet, Pierre [1 ]
机构
[1] CRIM, Montreal, PQ, Canada
来源
2013 8TH INTERNATIONAL WORKSHOP ON SYSTEMS, SIGNAL PROCESSING AND THEIR APPLICATIONS (WOSSPA) | 2013年
关键词
D O I
暂无
中图分类号
TM [电工技术]; TN [电子技术、通信技术];
学科分类号
0808 ; 0809 ;
摘要
This paper describes the French broadcast speech transcription system by CRIM for the ETAPE 2011 evaluation. The key elements in this recognizer include over 140,000-word dictionary, 478 hours of audio for training the acoustic models, feature-space MMI and boosted MMI discriminative training of the acoustic models, variable-frame-rate decoding with trigram language model, lattice rescoring with quadgram language model, soft penalty on silence models, confusion network decoding with minimum Bayes risk, and combining multiple recognizers with ROVER. Recognition enhancements after the ETAPE evaluation include discriminative training of the subspace Gaussian mixture models and lattice rescoring with neural net language models.
引用
收藏
页码:351 / 356
页数:6
相关论文
共 50 条
  • [41] Argument structure alternation in French children's speech
    Barrière, I
    Lorch, M
    Le Normand, MT
    NEW DIRECTIONS IN LANGUAGE DEVELOPMENT AND DISORDERS, 2000, : 139 - 147
  • [42] Towards an Automatic Speech-to-Text Transcription System: Amazigh Language
    Ouhnini, Ahmed
    Aksasse, Brahim
    Ouanan, Mohammed
    INTERNATIONAL JOURNAL OF ADVANCED COMPUTER SCIENCE AND APPLICATIONS, 2023, 14 (02) : 413 - 418
  • [43] Development of the CUHTK 2004 Mandarin conversational telephone speech transcription system
    Gales, MJF
    Jia, B
    Liu, X
    Sim, KC
    Woodland, P
    Yu, K
    2005 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH, AND SIGNAL PROCESSING, VOLS 1-5: SPEECH PROCESSING, 2005, : 841 - 844
  • [44] Using the lemmatization technique for phonetic transcription in text-to-speech system
    Kanis, J
    Müller, L
    TEXT, SPEECH AND DIALOGUE, PROCEEDINGS, 2004, 3206 : 355 - 361
  • [45] The ISL 2007 English Speech Transcription System for European Parliament Speeches
    Stueker, Sebastian
    Fuegen, Christian
    Kraft, Florian
    Woelfel, Matthias
    INTERSPEECH 2007: 8TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION, VOLS 1-4, 2007, : 673 - 676
  • [46] An efficient approach for the automated segmentation and transcription of the People's Speech corpus
    Biswas, Astik
    Boumadane, Abdelmoumene
    Peillon, Stephane
    Bleas, Gildas
    INTERSPEECH 2023, 2023, : 3939 - 3943
  • [47] French speech as dramatic action in Shakespeare's Henry V
    Walls, Alison
    LANGUAGE AND LITERATURE, 2013, 22 (02) : 119 - 131
  • [48] Respeak: A Voice-based, Crowd-powered Speech Transcription System
    Vashistha, Aditya
    Sethi, Pooja
    Anderson, Richard
    PROCEEDINGS OF THE 2017 ACM SIGCHI CONFERENCE ON HUMAN FACTORS IN COMPUTING SYSTEMS (CHI'17), 2017, : 1855 - 1866
  • [49] Modification of the Speech Feature Extraction Module for the Improvement of the System for Automatic Lectures Transcription
    Chaloupka, Josef
    Cerva, Petr
    Silovsky, Jan
    Zd'ansky, Jindrich
    Nouza, Jan
    PROCEEDINGS ELMAR-2012, 2012, : 223 - 226
  • [50] Development of the 2003 CU-HTK Conversational Telephone Speech transcription system
    Evermann, G
    Chan, HY
    Gales, MJF
    Hain, T
    Liu, X
    Mrva, D
    Wang, L
    Woodland, P
    2004 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH, AND SIGNAL PROCESSING, VOL I, PROCEEDINGS: SPEECH PROCESSING, 2004, : 249 - 252