--- pretty_name: Evaluation run of TheSkullery/Aurora-V2-DLEC dataset_summary: "Dataset automatically created during the evaluation run of model\ \ [TheSkullery/Aurora-V2-DLEC](https://huggingface.c
该数据集是在评估模型bardsai/jaskier-7b-dpo-v7.1时自动创建的,用于Open LLM Leaderboard。数据集包含63个配置,每个配置对应一个评估任务。数据集由1次运行生成,每次运行的结果作为一个特定的分割存储在配置中,分割名称使用运行的时间戳。train分割始终指向最新的结果。此外,results配置存储了所有运行的聚合结果,用于计算和显示Open LLM Lead
--- pretty_name: Evaluation run of fradinho/llama-mistral dataset_summary: "Dataset automatically created during the evaluation run of model\ \ [fradinho/llama-mistral](https://huggingface.co/fradin
--- pretty_name: Evaluation run of OpenBuddy/openbuddy-deepseek-67b-v15.2 dataset_summary: "Dataset automatically created during the evaluation run of model\ \ [OpenBuddy/openbuddy-deepseek-67b-v15.
--- pretty_name: Evaluation run of bsp-albz/llama2-13b-platypus-ckpt-1000 dataset_summary: "Dataset automatically created during the evaluation run of model\ \ [bsp-albz/llama2-13b-platypus-ckpt-100