相关数据集
GENIAC-team-haijima/yutohub-002-sft-data
--- dataset_info: features: - name: output dtype: string - name: input dtype: string - name: instruction dtype: string splits: - name: train num_bytes: 8959599 num_exam
Hugging Face2024-05-26 更新170
prestonfu/deepscaler-onpolicy-answers
该数据集包含问题、答案、解决方案、策略答案及其统计数据等字段。数据集分为训练集,共有40315个示例,总文件大小约为19.98MB。
Hugging Face2025-11-12 更新80
Arabic_summaries_batch36
该数据集包含三个字段:id,文本内容和摘要。文本内容字段包含了文本数据,而摘要字段可能包含对应文本的简短总结。数据集分为训练集,共有3600个示例。数据集的下载大小为9116920字节,总大小为19422235字节。
Hugging Face2025-03-06 更新50
open-llm-leaderboard-old/details_ChavyvAkvar__habib-v4
--- pretty_name: Evaluation run of ChavyvAkvar/habib-v4 dataset_summary: "Dataset automatically created during the evaluation run of model\ \ [ChavyvAkvar/habib-v4](https://huggingface.co/ChavyvAkva
Hugging Face2024-04-06 更新130
luca0621/multi-RLHF-processed-llama1B-dataset-with-1000-rewards
该数据集包含查询(query)、响应(response)和奖励(reward)三个特征,其中查询和响应为字符串类型,奖励为浮点数类型。数据集分为训练集和测试集,训练集包含8000个样本,测试集包含2000个样本。总下载大小为2276183字节,数据集总大小为7148705字节。
Hugging Face2024-11-30 更新70



