遇见数据集

ESpeech/ESpeech-tuchniyzhab

收藏
Hugging Face2025-08-25 更新2025-09-13 收录
官方服务:

资源简介:

Tuchniy Zhab YouTube音频数据集包含了从Tuchniy Zhab YouTube频道提取的306小时的音频片段及其对应的元数据。每个音频文件代表频道视频内容的一个片段,以44.1kHz采样率处理后保存为MP3格式。数据集适用于文本到语音、自动语音识别和语音质量评估任务,包含俄语语言的音频数据。

The Tuchniy Zhab YouTube Audio Dataset contains 306 hours of processed audio segments extracted from the Tuchniy Zhab YouTube channel along with corresponding metadata. Each audio file represents a segment from the channels videos and content, processed at 44.1kHz sample rate, and is saved in MP3 format. The dataset is suitable for tasks such as Text-to-Speech (TTS), Automatic Speech Recognition (ASR), and speech quality assessment, and includes audio data in the Russian language.

提供机构:
ESpeech
二维码
社区交流群
二维码
科研交流群
商业服务