euiyulsong/sciqa_train_ds_random_rejected
收藏资源简介:
--- dataset_info: features: - name: image dtype: image - name: question dtype: string - name: choices sequence: string - name: answer dtype: int8 - name: hint dtype: string - name: task dtype: string - name: grade dtype: string - name: subject dtype: string - name: topic dtype: string - name: category dtype: string - name: skill dtype: string - name: lecture dtype: string - name: solution dtype: string - name: prompt dtype: string - name: chosen dtype: string - name: rejected dtype: string splits: - name: train num_bytes: 425433188.182 num_examples: 12726 download_size: 411416631 dataset_size: 425433188.182 configs: - config_name: default data_files: - split: train path: data/train-* ---
This dataset includes various educational features such as images, questions, choices, answers, hints, and more, suitable for machine learning tasks in the education field. The dataset is divided into a training set with 12726 examples, with a total size of 425433188.182 bytes and a download size of 411416631 bytes.
数据集概述
数据集特征
- image:图像数据类型
- question:字符串数据类型
- choices:字符串序列数据类型
- answer:8位整数数据类型
- hint:字符串数据类型
- task:字符串数据类型
- grade:字符串数据类型
- subject:字符串数据类型
- topic:字符串数据类型
- category:字符串数据类型
- skill:字符串数据类型
- lecture:字符串数据类型
- solution:字符串数据类型
- prompt:字符串数据类型
- chosen:字符串数据类型
- rejected:字符串数据类型
数据集分割
- train:训练集,包含12726个样本,数据量约为425433188.182字节
数据集大小
- 下载大小:411416631字节
- 数据集大小:425433188.182字节
配置
- config_name: default
- data_files:
- split: train
- path: data/train-*



