CLUTRR/v1
收藏资源简介:
CLUTRR数据集是一个用于测试自然语言理解系统在系统性泛化和归纳推理能力方面的诊断基准。该数据集包含大量半合成的家庭故事,任务是推断故事中未明确提及的两个家庭成员之间的关系。数据集的结构包括多个配置,每个配置包含故事、查询、目标关系等信息。数据集的划分包括训练、验证和测试集,每个划分的实例数量也有所不同。
The CLUTRR dataset is a diagnostic benchmark developed to test the systematic generalization and inductive reasoning abilities of natural language understanding systems. It includes a large corpus of semi-synthetic family stories, with the core task being to infer the relational ties between two family members that are not explicitly mentioned in the given story. The dataset is structured with multiple configurations, each containing relevant information such as the story, query prompt, and target relational category. Additionally, the dataset is partitioned into training, validation, and test subsets, with varying numbers of instances across each split.
数据集概述
数据集名称
CLUTRR (Compositional Language Understanding and Text-based Relational Reasoning)
数据集描述
- 目的: 测试自然语言理解(NLU)系统的系统性泛化和归纳推理能力。
- 内容: 包含大量涉及假设家庭的半合成故事,任务是推断故事中未明确提及的两个家庭成员之间的关系。
数据集任务
- 目标: 确定两个家庭成员之间的正确关系。
- 关系类型: 包括“aunt”, “son-in-law”, “grandfather”等21种关系,每种关系有对应的编号。
数据集结构
- 配置: 数据集包含14种配置,每种配置包括id, story, query, target, target_text等字段。
- 实例示例: 包括故事文本、查询关系、目标关系及其文本描述、逻辑规则等详细信息。
数据分割
- 分割名称: 包括gen_train23_test2to10, gen_train234_test2to10等。
- 分割详情: 每个分割包含训练、验证和测试集,详细记录了每个分割中的实例数量。
多语言性
- 语言: 单语(英语)
数据集大小
- 规模: 10K<n<100K
许可证
- 许可证类型: 未知




