AILab-CVC/SEED-Bench-2-plus
收藏资源简介:
--- license: cc-by-nc-4.0 task_categories: - visual-question-answering language: - en pretty_name: SEED-Bench-2-Plus size_categories: - 1K<n<10K --- # SEED-Bench-2-Plus Card ## Benchmark details **Benchmark type:** SEED-Bench-2-Plus is a large-scale benchmark to evaluate Multimodal Large Language Models (MLLMs). It consists of 2.3K multiple-choice questions with precise human annotations, spanning three broad categories: Charts, Maps, and Webs, each of which covers a wide spectrum of text-rich scenarios in the real world. **Benchmark date:** SEED-Bench-2-Plus was collected in April 2024. **Paper or resources for more information:** https://github.com/AILab-CVC/SEED-Bench **License:** Attribution-NonCommercial 4.0 International. It should abide by the policy of OpenAI: https://openai.com/policies/terms-of-use. For the images of SEED-Bench-2-plus, we use data from the internet under CC-BY licenses. Please contact us if you believe any data infringes upon your rights, and we will remove it. **Where to send questions or comments about the benchmark:** https://github.com/AILab-CVC/SEED-Bench/issues ## Intended use **Primary intended uses:** The primary use of SEED-Bench-2-Plus is evaluate Multimodal Large Language Models on text-rich visual understanding. **Primary intended users:** The primary intended users of the Benchmark are researchers and hobbyists in computer vision, natural language processing, machine learning, and artificial intelligence.
--- 许可证:CC-BY-NC-4.0 任务类别: - 视觉问答(visual-question-answering) 语言: - 英语 展示名称:SEED-Bench-2-Plus 规模类别: - 1000 < 样本量 < 10000 --- # SEED-Bench-2-Plus 数据集卡片 ## 基准详情 **基准类型:** SEED-Bench-2-Plus 是一款用于评估多模态大语言模型(Multimodal Large Language Models, MLLMs)的大规模基准测试集。该基准集包含2300道经人工精准标注的多项选择题,涵盖图表(Charts)、地图(Maps)与网页(Webs)三大类别,每一类均覆盖现实世界中各类富含文本的场景。 **基准采集时间:** SEED-Bench-2-Plus 于2024年4月完成数据采集。 **详细信息参阅资源:** https://github.com/AILab-CVC/SEED-Bench **许可证:** 采用署名-非商业性使用4.0国际许可协议(Attribution-NonCommercial 4.0 International),同时需遵守OpenAI相关政策:https://openai.com/policies/terms-of-use。 SEED-Bench-2-Plus 所使用的图像数据均来自互联网,采用CC-BY许可协议。若您认为本基准集中的任何数据侵犯了您的合法权益,请联系我们,我们将立即予以移除。 **基准相关问题反馈渠道:** 请前往:https://github.com/AILab-CVC/SEED-Bench/issues 提交相关问题或意见。 ## 预期用途 **主要预期用途:** SEED-Bench-2-Plus 的核心用途为评估多模态大语言模型在富含文本的视觉理解任务上的性能。 **主要目标用户:** 本基准集的主要目标用户为计算机视觉、自然语言处理、机器学习以及人工智能领域的研究人员与爱好者。
SEED-Bench-2-Plus 数据集概述
基本信息
- 许可证: cc-by-nc-4.0
- 任务类别: 视觉问答
- 语言: 英语
- 数据集大小: 1K<n<10K
数据集详情
- 类型: SEED-Bench-2-Plus 是一个大规模基准测试,用于评估多模态大型语言模型(MLLMs)。
- 包含内容: 包含2.3K个多选题,涵盖图表、地图和网络三大类别,涉及现实世界中丰富的文本场景。
- 收集时间: 2024年4月
- 更多信息资源: SEED-Bench GitHub链接
使用目的
- 主要用途: 评估多模态大型语言模型在文本丰富的视觉理解能力。
- 目标用户: 计算机视觉、自然语言处理、机器学习和人工智能领域的研究人员及爱好者。




