遇见数据集

BIG-Bench Evaluation Benchmark

收藏
arXiv2025-09-30 收录
官方服务:

资源简介:

该数据集是一个综合性基准数据集,涵盖了包括图像分类、神经机器翻译、语言建模以及其它语言相关评估在内的90个评测任务。具体来说,它包含了72个图像分类任务、5个神经机器翻译任务、5个语言建模任务以及10个其他语言相关评估任务,为研究人员提供了一个全面的评测基准。

This dataset is a comprehensive benchmark dataset covering 90 evaluation tasks, including image classification, neural machine translation, language modeling, and other language-related evaluations. Specifically, it comprises 72 image classification tasks, 5 neural machine translation tasks, 5 language modeling tasks, and 10 other language-related evaluation tasks, providing researchers with a comprehensive evaluation benchmark.

提供机构:
Google Research
二维码
社区交流群
二维码
科研交流群
商业服务