Automatic Corpus Selecting Algorithm Based on Triphone Models
DOI:
Author:
Affiliation:

Clc Number:

Fund Project:

  • Article
  • |
  • Figures
  • |
  • Metrics
  • |
  • Reference
  • |
  • Related
  • |
  • Cited by
  • |
  • Materials
  • |
  • Comments
    Abstract:

    In speech recognition,the selection of training corpus for robust acoustic modeling which can cover almost all phone phenomena is very important.Traditionally,corpus is selected manually first,and then tested and supplemented,which can't provide sufficient coverage of samples for various statistical modeling methods.An algorithm for automatically selecting the training samples from large-scale text corpus is developed in this paper.This algorithm can not only cover almost all phone phenomena but also ensure to include ideal samples of triphones or class-triphones and ensure enough data for training,which makes it possible to train acoustic model reliably.

    Reference
    Related
    Cited by
Get Citation

吴华,徐波,黄泰翼.基于三音子模型的语料自动选择算法.软件学报,2000,11(2):271-276

Copy
Share
Article Metrics
  • Abstract:
  • PDF:
  • HTML:
  • Cited by:
History
  • Received:January 07,1999
  • Revised:March 18,1999
  • Adopted:
  • Online:
  • Published:
You are the firstVisitors
Copyright: Institute of Software, Chinese Academy of Sciences Beijing ICP No. 05046678-4
Address:4# South Fourth Street, Zhong Guan Cun, Beijing 100190,Postal Code:100190
Phone:010-62562563 Fax:010-62562533 Email:jos@iscas.ac.cn
Technical Support:Beijing Qinyun Technology Development Co., Ltd.

Beijing Public Network Security No. 11040202500063