DISCRIMINATIVE SEGMENTAL CASCADES FOR FEATURE-RICH PHONE RECOGNITION

被引:0
|
作者
Tang, Hao [1 ]
Wang, Weiran [1 ]
Gimpel, Kevin [1 ]
Livescu, Karen [1 ]
机构
[1] Toyota Technol Inst, Chicago, IL 60637 USA
关键词
segmental conditional random field; structured prediction cascades; phone recognition; segment neural network; beam search;
D O I
暂无
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Discriminative segmental models, such as segmental conditional random fields (SCRFs) and segmental structured support vector machines (SSVMs), have had success in speech recognition via both lattice rescoring and first-pass decoding. However, such models suffer from slow decoding, hampering the use of computationally expensive features, such as segment neural networks or other high-order features. A typical solution is to use approximate decoding, either by beam pruning in a single pass or by beam pruning to generate a lattice followed by a second pass. In this work, we study discriminative segmental models trained with a hinge loss (i.e., segmental structured SVMs). We show that beam search is not suitable for learning rescoring models in this approach, though it gives good approximate decoding performance when the model is already well-trained. Instead, we consider an approach inspired by structured prediction cascades, which use max-marginal pruning to generate lattices. We obtain a high-accuracy phonetic recognition system with several expensive feature types: a segment neural network, a second-order language model, and second-order phone boundary features.
引用
收藏
页码:561 / 568
页数:8
相关论文
共 50 条
  • [31] Feature-rich Interworking Architecture for Mobile Traffic Offloading
    Beyene, Asrat Mulatu
    Serra, Jordi Casademont
    Shiferaw, Yalemzewd Negash
    PROCEEDINGS OF THE IEEE INTERNATIONAL CONFERENCE ON COMPUTING NETWORKING AND INFORMATICS (ICCNI 2017), 2017,
  • [32] Feature-rich flash memories deliver high density
    Bursky, D
    ELECTRONIC DESIGN, 1999, 47 (16) : 67 - +
  • [33] Feature-Rich Electronic Properties of Sliding Bilayer Germanene
    Liu, Hsin-Yi
    Wu, Jhao-Ying
    ACS OMEGA, 2022, 7 (46): : 42304 - 42312
  • [34] Parcels: A fast and feature-rich binary deployment technology
    Miranda, E
    Leibs, D
    Wuyts, R
    COMPUTER LANGUAGES SYSTEMS & STRUCTURES, 2005, 31 (3-4) : 165 - 181
  • [35] Feature-rich distance-based terrain synthesis
    Brennan Rusnell
    David Mould
    Mark Eramian
    The Visual Computer, 2009, 25 : 573 - 579
  • [36] Regularized discriminative segmental feature transform method
    Chen B.
    Zhang L.
    Qu D.
    Li B.
    Xi'an Dianzi Keji Daxue Xuebao/Journal of Xidian University, 2016, 43 (02): : 102 - 107
  • [37] Feature-rich Regular Expression Matching Accelerator for Text Analytics
    Kubilay Atasu
    Journal of Signal Processing Systems, 2016, 85 : 355 - 371
  • [38] Rendering of Feature-Rich Dynamically Changing Volumetric Datasets on GPU
    Schreiber, Martin
    Atanasov, Atanas
    Neumann, Philipp
    Bungartz, Hans-Joachim
    2014 INTERNATIONAL CONFERENCE ON COMPUTATIONAL SCIENCE, 2014, 29 : 648 - 658
  • [39] Intrinsic Plagiarism Detection with Feature-Rich Imbalanced Dataset Learning
    Polydouri, Andrianna
    Siolas, Georgios
    Stafylopatis, Andreas
    ENGINEERING APPLICATIONS OF NEURAL NETWORKS, EANN 2017, 2017, 744 : 99 - 110
  • [40] Feature-rich Regular Expression Matching Accelerator for Text Analytics
    Atasu, Kubilay
    JOURNAL OF SIGNAL PROCESSING SYSTEMS FOR SIGNAL IMAGE AND VIDEO TECHNOLOGY, 2016, 85 (03): : 355 - 371