遇见数据集

NeutrinoPit/OpenSubtitles2024-en-ar-batch53

收藏
Hugging Face2025-03-04 更新2025-04-12 收录
官方服务:

资源简介:

该数据集包含两种语言的文本数据:英语(en)和阿拉伯语(ar)。它包含一个训练集,共有100万个示例,文件大小为约100MB。数据集的具体内容未在README中描述,但可以推断它是一个用于语言处理的文本数据集。

The dataset consists of text data in two languages: English (en) and Arabic (ar). It includes a training set with 1,000,000 examples, totaling approximately 100MB in size. The specific content of the dataset is not described in the README, but it can be inferred to be a text dataset for language processing.

提供机构:
NeutrinoPit
二维码
社区交流群
二维码
科研交流群
商业服务