遇见数据集

West Point Russian Speech

收藏
Mendeley Data2024-01-31 更新2024-06-28 收录
官方服务:

资源简介:

Introduction West Point Russian Speech was developed at the Department of Foreign Languages (DFL) and the Center for Technology Enhanced Language Learning (CTELL) at the United States Military Academy at West Point. The purpose of the corpus is to provide a set of recordings for the training and development of speaker-independent speech recognition systems for use by West Point cadets enrolled in the Russian language program. Data The corpus consists of 4,181 speech files in SPHERE format, totalling approximately four hours of speech. Approximately 2,290 files are from native informants and 1,891 are from non-native informants. The following tables show the breakdown of corpus content in terms of male, female, native and non-native speakers. Number of speakers: male female total native 13 16 29 non-native 16 10 26 totals 29 26 55 Number of speech files: male female total native 1027 1263 2290 non-native 1103 788 1891 totals 2130 2050 4181 The speech data was collected using laptop computers running Windows NT. Recordings were captured at a sampling rate of 16-bit at 22,050 Hz pcm using a Shure SM10A microphone and a RANE Model MS1 pre-amplifier. A visual display of the sentence, along with a digital recording of the sentence as read by a native speaker, was presented. The informant pressed the Enter key to record the utterance. The informant's recording was played back for review and the utterance was re-recorded if necessary. The collection script consists of 96 sentences with a total of 528 tokens and 351 types. Each waveform file has a monophone and word level master label file transcription in HTK-format. A concatenated version of the master label files at both the word level and the phone level is provided. The lexicon contains 690 distinct orthographic word forms, including all words found in the collection script. Samples Please view the following samples: Female Speaker (S31) Male Speaker (S08) Phone Level Transcript Word Level Transcript Updates There are no updates available at this time. Portions © 2003 United States Military Academy, © 2003 Trustees of the University of Pennsylvania

引言 本数据集为西点俄语语音语料库(West Point Russian Speech),由美国西点军校(United States Military Academy at West Point)外语系(Department of Foreign Languages, DFL)与技术增强语言学习中心(Center for Technology Enhanced Language Learning, CTELL)联合开发。本语料库旨在为面向西点军校俄语课程学员的与说话人无关的语音识别系统的训练与开发提供标准化录音数据集。 数据概况 该语料库包含4181个SPHERE格式的语音文件,总时长约4小时。其中约2290个文件来自母语使用者,1891个文件来自非母语使用者。按性别与母语属性划分的语料库内容统计如下: 1. 说话人数量: 母语使用者:男性13人,女性16人,总计29人 非母语使用者:男性16人,女性10人,总计26人 总计:男性29人,女性26人,总计55人 2. 语音文件数量: 母语使用者:男性1027个,女性1263个,总计2290个 非母语使用者:男性1103个,女性788个,总计1891个 总计:男性2130个,女性2050个,总计4181个 语音数据采集说明 语音数据采用运行Windows NT系统的笔记本电脑完成采集。录音采用Shure SM10A麦克风与RANE Model MS1前置放大器,以22050Hz的采样率采集16位脉冲编码调制(PCM)音频。采集过程中,系统会同步展示目标句子的可视化文本,并附带一名母语使用者朗读该句子的示范录音。受试人员按下回车键即可开始录制语音,录制完成后系统会回放录音供审核,若存在质量问题则可重新录制。 标注与词典说明 本次采集的脚本包含96个句子,总计包含528个Token与351个词型。每个波形文件均附带HTK格式的单音素级与词级主标注转录文件。此外,还提供了词级与音素级主标注文件的拼接版本。本语料库的词典包含690种不同的正词法词形,涵盖采集脚本中的全部词汇。 示例 请查看以下示例: 女性说话人(S31)、男性说话人(S08) 音素级转录文本 词级转录文本 更新说明 目前暂无可用更新。 版权声明 部分内容 © 2003 美国西点军校,© 2003 宾夕法尼亚大学理事会

创建时间:
2024-01-31
二维码
社区交流群
二维码
科研交流群
商业服务