遇见数据集

gowitheflow/supervised-multilingual

收藏
Hugging Face2024-07-26 更新2024-12-14 收录
官方服务:

资源简介:

该数据集包含两个文本字段,分别为sentence1和sentence2,主要用于文本对的分析或比较。数据集包含一个训练分割(train),共有19,624,898个样本,总数据大小约为6.41GB,下载大小约为4.77GB。

This dataset includes two text fields, sentence1 and sentence2, primarily used for the analysis or comparison of text pairs. The dataset contains one training split (train) with a total of 19,624,898 samples, a total data size of approximately 6.41GB, and a download size of approximately 4.77GB.

提供机构:
gowitheflow
二维码
社区交流群
二维码
科研交流群
商业服务