遇见数据集

ivanjaenm/ot-dataset_bins50_size500k_prec3_zero-supress

收藏
Hugging Face2025-09-22 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含了五个特征:源距离(source_dist)、目标距离(target_dist)、输入索引(input_idx)、总输入(total_input)和输出映射稀疏矩阵(output_map_sparse)。数据集分为训练集和测试集,其中训练集包含2000万个示例,测试集包含500万个示例。数据集的总大小为41,360,699,530字节。

The dataset includes five features: source distance (source_dist), target distance (target_dist), input index (input_idx), total input (total_input), and output map sparse matrix (output_map_sparse). The dataset is split into a training set and a test set, with the training set containing 20 million examples and the test set containing 5 million examples. The total size of the dataset is 41,360,699,530 bytes.

提供机构:
ivanjaenm
二维码
社区交流群
二维码
科研交流群
商业服务