CRIM'S FRENCH SPEECH TRANSCRIPTION SYSTEM FOR ETAPE 2011

被引:0
|
作者
Gupta, Vishwa [1 ]
Boulianne, Gilles [1 ]
Osterrath, Frederic [1 ]
Ouellet, Pierre [1 ]
机构
[1] CRIM, Montreal, PQ, Canada
关键词
D O I
暂无
中图分类号
TM [电工技术]; TN [电子技术、通信技术];
学科分类号
0808 ; 0809 ;
摘要
This paper describes the French broadcast speech transcription system by CRIM for the ETAPE 2011 evaluation. The key elements in this recognizer include over 140,000-word dictionary, 478 hours of audio for training the acoustic models, feature-space MMI and boosted MMI discriminative training of the acoustic models, variable-frame-rate decoding with trigram language model, lattice rescoring with quadgram language model, soft penalty on silence models, confusion network decoding with minimum Bayes risk, and combining multiple recognizers with ROVER. Recognition enhancements after the ETAPE evaluation include discriminative training of the subspace Gaussian mixture models and lattice rescoring with neural net language models.
引用
收藏
页码:351 / 356
页数:6
相关论文
共 50 条
  • [1] CRIM's Speech Transcription and Call Sign Detection System for the ATC Airbus Challenge task
    Gupta, Vishwa
    Rebout, Lise
    Boulianne, Gilles
    Menard, Pierre-Andre
    Alam, Jahangir
    INTERSPEECH 2019, 2019, : 3018 - 3022
  • [2] The ETAPE corpus for the evaluation of speech-based TV content processing in the French language
    Gravier, Guillaume
    Adda, Gilles
    Paulsson, Niklas
    Carre, Matthieu
    Giraudel, Aude
    Galibert, Olivier
    LREC 2012 - EIGHTH INTERNATIONAL CONFERENCE ON LANGUAGE RESOURCES AND EVALUATION, 2012, : 114 - 118
  • [3] CRIM's System for the MGB-3 English Multi-Genre Broadcast Media Transcription
    Gupta, Vishwa
    Boulianne, Gilles
    19TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2018), VOLS 1-6: SPEECH RESEARCH FOR EMERGING MARKETS IN MULTILINGUAL SOCIETIES, 2018, : 2653 - 2657
  • [4] CRIM's Speech Recognition System for OpenASR21 Evaluation with Conformer and Voice Activity Detector Embeddings
    Gupta, Vishwa
    Boulianne, Gilles
    SPEECH AND COMPUTER, SPECOM 2022, 2022, 13721 : 238 - 251
  • [5] Transcription of Children's Speech
    Wren, Yvonne
    McLeod, Sharynne
    Verdon, Sarah
    FOLIA PHONIATRICA ET LOGOPAEDICA, 2020, 72 (02) : 73 - 74
  • [6] A mandarin lecture speech transcription system for speech summarization
    Chan, Ho Yin
    Zhang, Justin Jian
    Fung, Pascale
    Cao, Lu
    2007 IEEE WORKSHOP ON AUTOMATIC SPEECH RECOGNITION AND UNDERSTANDING, VOLS 1 AND 2, 2007, : 467 - 471
  • [7] The AMI system for the transcription of speech in meetings
    Hain, Thomas
    Burget, Lukas
    Dines, John
    Garau, Giulia
    Karafiat, Martin
    Lincoln, Mike
    Vepa, Jithendra
    Wan, Vincent
    2007 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH, AND SIGNAL PROCESSING, VOL IV, PTS 1-3, 2007, : 357 - +
  • [8] The IBM BOLT Speech Transcription System
    Thomas, Samuel
    Saon, George
    Kuo, Hong-Kwang
    Mangu, Lidia
    16TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2015), VOLS 1-5, 2015, : 3150 - 3153
  • [9] LIUM ASR System for ETAPE French Evaluation Campaign: Experiments on System Combination Using Open-Source Recognizers
    Bougares, Fethi
    Deleglise, Paul
    Esteve, Yannick
    Rouvier, Mickael
    TEXT, SPEECH, AND DIALOGUE, TSD 2013, 2013, 8082 : 319 - 326
  • [10] The IBM mandarin broadcast speech transcription system
    Chu, Stephen M.
    Kuo, Hong-kwang
    Liu, Yi Y.
    Qin, Yong
    Shi, Qin
    Zweig, Geoffrey
    2007 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH, AND SIGNAL PROCESSING, VOL II, PTS 1-3, 2007, : 345 - +