遇见数据集

sumanpaudel1997/hindi_tweet_sampled_dataset

收藏
Hugging Face2024-10-19 更新2024-12-14 收录
官方服务:

资源简介:

该数据集主要包含以印地语(hi)为主的文本数据,遵循Apache 2.0许可证。数据集的特征包括一个名为text的字符串类型字段。数据集被分割为训练集,包含10,183,176个示例,总大小为2,841,618,218字节,下载大小为1,297,427,903字节。

This dataset primarily contains text data in Hindi (hi), licensed under Apache 2.0. The dataset features include a string-type field named text. The dataset is split into a training set, containing 10,183,176 examples, with a total size of 2,841,618,218 bytes and a download size of 1,297,427,903 bytes.

提供机构:
sumanpaudel1997
二维码
社区交流群
二维码
科研交流群
商业服务