CRIM'S FRENCH SPEECH TRANSCRIPTION SYSTEM FOR ETAPE 2011

被引：0

作者：

Gupta, Vishwa ^{[1
]}

Boulianne, Gilles ^{[1
]}

Osterrath, Frederic ^{[1
]}

Ouellet, Pierre ^{[1
]}

机构：

[1] CRIM, Montreal, PQ, Canada

来源：

2013 8TH INTERNATIONAL WORKSHOP ON SYSTEMS, SIGNAL PROCESSING AND THEIR APPLICATIONS (WOSSPA) | 2013年

关键词：

D O I：

暂无

中图分类号：

TM [电工技术]; TN [电子技术、通信技术];

学科分类号：

0808 ; 0809 ;

摘要：

This paper describes the French broadcast speech transcription system by CRIM for the ETAPE 2011 evaluation. The key elements in this recognizer include over 140,000-word dictionary, 478 hours of audio for training the acoustic models, feature-space MMI and boosted MMI discriminative training of the acoustic models, variable-frame-rate decoding with trigram language model, lattice rescoring with quadgram language model, soft penalty on silence models, confusion network decoding with minimum Bayes risk, and combining multiple recognizers with ROVER. Recognition enhancements after the ETAPE evaluation include discriminative training of the subspace Gaussian mixture models and lattice rescoring with neural net language models.

引用

页码：351 / 356

页数：6

共 50 条

[1] CRIM's Speech Transcription and Call Sign Detection System for the ATC Airbus Challenge task
Gupta, Vishwa
Rebout, Lise
Boulianne, Gilles
Menard, Pierre-Andre
Alam, Jahangir
INTERSPEECH 2019, 2019, : 3018 - 3022
[2] The ETAPE corpus for the evaluation of speech-based TV content processing in the French language
Gravier, Guillaume
Adda, Gilles
Paulsson, Niklas
Carre, Matthieu
Giraudel, Aude
Galibert, Olivier
LREC 2012 - EIGHTH INTERNATIONAL CONFERENCE ON LANGUAGE RESOURCES AND EVALUATION, 2012, : 114 - 118
[3] CRIM's System for the MGB-3 English Multi-Genre Broadcast Media Transcription
Gupta, Vishwa
Boulianne, Gilles
19TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2018), VOLS 1-6: SPEECH RESEARCH FOR EMERGING MARKETS IN MULTILINGUAL SOCIETIES, 2018, : 2653 - 2657
[4] CRIM's Speech Recognition System for OpenASR21 Evaluation with Conformer and Voice Activity Detector Embeddings
Gupta, Vishwa
Boulianne, Gilles
SPEECH AND COMPUTER, SPECOM 2022, 2022, 13721 : 238 - 251
[5] Transcription of Children's Speech
Wren, Yvonne
McLeod, Sharynne
Verdon, Sarah
FOLIA PHONIATRICA ET LOGOPAEDICA, 2020, 72 (02) : 73 - 74
[6] A mandarin lecture speech transcription system for speech summarization
Chan, Ho Yin
Zhang, Justin Jian
Fung, Pascale
Cao, Lu
2007 IEEE WORKSHOP ON AUTOMATIC SPEECH RECOGNITION AND UNDERSTANDING, VOLS 1 AND 2, 2007, : 467 - 471
[7] The AMI system for the transcription of speech in meetings
Hain, Thomas
Burget, Lukas
Dines, John
Garau, Giulia
Karafiat, Martin
Lincoln, Mike
Vepa, Jithendra
Wan, Vincent
2007 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH, AND SIGNAL PROCESSING, VOL IV, PTS 1-3, 2007, : 357 - +
[8] The IBM BOLT Speech Transcription System
Thomas, Samuel
Saon, George
Kuo, Hong-Kwang
Mangu, Lidia
16TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2015), VOLS 1-5, 2015, : 3150 - 3153
[9] LIUM ASR System for ETAPE French Evaluation Campaign: Experiments on System Combination Using Open-Source Recognizers
Bougares, Fethi
Deleglise, Paul
Esteve, Yannick
Rouvier, Mickael
TEXT, SPEECH, AND DIALOGUE, TSD 2013, 2013, 8082 : 319 - 326
[10] The IBM mandarin broadcast speech transcription system
Chu, Stephen M.
Kuo, Hong-kwang
Liu, Yi Y.
Qin, Yong
Shi, Qin
Zweig, Geoffrey
2007 IEEE INTERNATIONAL CONFERENCE ON ACOUSTICS, SPEECH, AND SIGNAL PROCESSING, VOL II, PTS 1-3, 2007, : 345 - +

← 1 2 3 4 5 →