登录后查看消息通知
搜索
常见问题
消息
登录
首页
/
数据集
/
EAR Benchmark
EAR Benchmark
收藏
Figshare
2025-06-20 更新
2026-04-08 收录
实体消歧
视觉语言模型
数据链接:
https://figshare.com/articles/dataset/EAR_Benchmark/29368739/1
数据链接
链接失效反馈
官方服务:
问题咨询
购买咨询
在线客服
NEW
资源简介:
EARBench: Evaluating the Entity Ambiguity Resolution Ability of LLMs
应用场景:
提供机构:
Chen
创建时间:
2025-06-20
相关数据集
sled-umich/ROPE
视觉语言模型
多对象幻觉
ROPE数据集旨在通过利用现有的全景分割数据集(如MSCOCO-Panoptic和ADE20K)来评估和分析多目标幻觉现象。数据集包含多样对象及其实例级语义注释,分为训练和验证两部分,每部分根据图像中对象类别的分布进一步细分为同质、异质、野外和对抗性子集。文件结构包括图像和JSON文件,图像分为原始图像和带有边界框的可视化图像。JSON文件结构详细描述了图像文件夹、文件名、来源、尺寸、分割信息、对
Hugging Face
2024-07-19 更新
53
0
Towards More Dependable Specifications: An Empirical Study Exploring the Synergy of Traditional and LLM-Based Repair Approaches
软件规范修复
视觉语言模型
Declarative specification languages like Alloy are critical for modeling and verifying complex software systems, yet repairing these specifications remains a significant challenge for ensuring softwar
DataCite Commons
2025-04-01 更新
8
0
methane69/finetuning_LLaVA
视觉语言模型
模型微调
--- dataset_info: features: - name: image dtype: image - name: caption dtype: string splits: - name: train num_bytes: 8426655.0 num_examples: 138 - name: test num_bytes
Hugging Face
2024-03-08 更新
6
0
CountBenchQA
视觉语言模型
计数能力
该数据集是为评估视觉语言模型中的计数能力而引入的,基于PaliGemma项目。数据集包含491张图像,这些图像来自原始的CountBench数据集,但由于某些原始URL无法访问,因此只保留了这些图像。每张图像都配有一个文本描述和一个关于图像中对象数量的手动生成的问题。数据集的特征包括图像、文本、问题和数量。数据集分为一个测试集,包含491个样本。
Hugging Face
2024-10-21 更新
25
0
swap-uniba/EXAMS-V_IT
视觉语言模型
多模态评估
--- dataset_info: features: - name: image dtype: image - name: id dtype: int64 - name: question dtype: string - name: answer dtype: string splits: - name: test num_by
Hugging Face
2024-10-17 更新
7
0
© 2023-2026 上海数据发展科技有限责任公司 版权所有
沪ICP备17003045号-15
沪公网安备31010402336585号
热门搜索
社区交流群
科研交流群
商业服务
数据资源
寻源服务
数据采集
标注服务
数据产品
代理销售
数据领域
凭证登记
数据产品
介绍推广