ManaTTS-Persian-Speech-Dataset
收藏资源简介:
ManaTTS是最大的公开可访问的单说话者波斯语语料库,包含大约86小时的音频,采样率为44.1 kHz。该数据集在开放的CC-0许可下发布,适用于教育和商业用途。这个数据集是波斯语的综合语音数据集,从Nasl-e-Mana杂志收集,涵盖了广泛的主题和领域,适合训练高质量的文本到语音模型。数据集附带了一个完全透明、开源的数据收集和处理管道,包括音频分割和强制对齐的工具。
ManaTTS is the largest publicly accessible single-speaker Persian speech corpus, containing approximately 86 hours of audio with a sampling rate of 44.1 kHz. Released under the open CC-0 license, this dataset is available for both educational and commercial use. As a comprehensive Persian speech dataset collected from Nasl-e-Mana Magazine, it covers a wide range of topics and domains, making it suitable for training high-quality text-to-speech models. The dataset also comes with a fully transparent, open-source data collection and processing pipeline, including tools for audio segmentation and forced alignment.
ManaTTS-Persian-Speech-Dataset
概述
- 语言: 波斯语
- 时长: 约86小时
- 采样率: 44.1 kHz
- 许可: CC-0 1.0(允许教育和商业用途)
- 来源: Nasl-e-Mana 杂志
- 适用场景: 训练高质量的文本到语音模型
数据集
- 下载链接: ManaTTS数据集
- 样本数据: 样本数据目录
数据采集
- 原始数据: 从Nasl-e-Mana杂志网站爬取
- 爬虫脚本: Google Colab链接
处理流程
- 流程图: resources/image.png
- Jupyter Notebook: Google Colab链接
训练模型
贡献
- 贡献方式: 欢迎提交问题或拉取请求
许可
- 数据集: CC-0 1.0
- 处理流程: MIT 许可
伦理使用
- 使用目的: 仅限于研究和开发
- 禁止行为: 禁止语音模仿、身份盗窃或欺诈活动
致谢
- 感谢对象: Nasl-e-Mana 杂志




