Project PIAF: 原生法语问答数据集
收藏资源简介:
Project PIAF是一个专注于收集原生法语问答数据的项目,由法国的研究机构reciTAL和Etalab共同创建。该数据集包含3835个问答对,主要用于评估非英语语言的下游任务,如问答系统。数据集的创建过程采用了参与式方法,通过组织多场现场标注活动(annotathons)来收集数据,参与者包括志愿者和PIAF团队成员。数据集的应用领域主要集中在自然语言处理和人工智能领域,旨在解决法语环境下问答系统的数据稀缺问题。
Project PIAF is a project dedicated to collecting native French question-answering data, co-established by two French research institutions, reciTAL and Etalab. This dataset contains 3,835 question-answering pairs, and is primarily utilized for evaluating downstream tasks in non-English languages, such as question-answering systems. The dataset was developed using a participatory approach, with data collected via multiple on-site annotation events (annotathons) involving volunteers and members of the PIAF team. Its application fields mainly focus on natural language processing (NLP) and artificial intelligence (AI), aiming to address the data scarcity problem of question-answering systems in the French language context.

- 1Project PIAF: Building a Native French Question-Answering DatasetreciTAL, 巴黎 (法国) ‡Etalab, DINUM, 总理办公室, 巴黎 (法国) · 2020年



