Abstract
The Taiwan Mandarin Radio Speech Corpus contains 300 (and growing) hours of high-quality recordings selected from Taiwan's National Education Radio (NER) archive. The corpus features speech (of various speaking styles, produced by hundreds of speakers) and their corresponding transcriptions (automatically transcribed and manually corrected) and annotations, which are suitable for speech and language research. In this paper, we report the progress of the corpus development and especially show the experimental results of audio event detection/segmentation and semi-supervised acoustic model training on this corpus.
| Original language | English |
|---|---|
| Title of host publication | 2017 20th Conference of the Oriental Chapter of International Committee for Coordination and Standardization of Speech Databases and Assessment Techniques, O-COCOSDA 2017 |
| Publisher | Institute of Electrical and Electronics Engineers Inc. |
| Pages | 1-6 |
| Number of pages | 6 |
| ISBN (Electronic) | 9781538633335 |
| DOIs | |
| State | Published - 13 Jun 2018 |
| Event | 20th Conference of the Oriental Chapter of International Committee for Coordination and Standardization of Speech Databases and Assessment Techniques, O-COCOSDA 2017 - Seoul, Korea, Republic of Duration: 1 Nov 2017 → 3 Nov 2017 |
Publication series
| Name | 2017 20th Conference of the Oriental Chapter of International Committee for Coordination and Standardization of Speech Databases and Assessment Techniques, O-COCOSDA 2017 |
|---|
Conference
| Conference | 20th Conference of the Oriental Chapter of International Committee for Coordination and Standardization of Speech Databases and Assessment Techniques, O-COCOSDA 2017 |
|---|---|
| Country/Territory | Korea, Republic of |
| City | Seoul |
| Period | 1/11/17 → 3/11/17 |
Bibliographical note
Publisher Copyright:© 2017 IEEE.
Keywords
- Mandarin speech corpus
- audio event detection
- semi-supervised training
Fingerprint
Dive into the research topics of 'A progress report of the Taiwan Mandarin radio speech corpus project'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver