相关数据集
ferrazzipietro/noLoraLS_Llama-2-7b-hf_adapters_en.layer1_NoQuant_2_0.0002_3EpochsLast
--- dataset_info: features: - name: sentence dtype: string - name: entities list: - name: id dtype: string - name: offsets sequence: int64 - name: role dtyp
Hugging Face2024-05-18 更新120
nqv2291/eval-vi-ner-VinAI-COVID_19_NER
该数据集包含三个特征:text(文本)、entity_type(实体类型)和label(标签)。数据集仅包含一个测试集(test),该测试集包含7075个样本,文件大小为1745459字节。下载大小为409619字节,数据集总大小为1745459字节。数据集的配置文件名为default,数据文件路径为data/test-*。
Hugging Face2024-06-27 更新90
disi-unibo-nlp/biored
这是一个用于自然语言处理的文本数据集,包含分词(tokens)和命名实体识别标签(ner_tags)。数据集分为训练集、验证集和测试集,共计包含5517个示例。训练集包含4343个示例,大小为1963620字节;验证集包含1127个示例,大小为521918字节;测试集包含1097个示例,大小为504330字节。数据集的总下载大小为624368字节,存储大小为2989868字节。
Hugging Face2025-07-04 更新40
NERSkill.Id
NERSkill.Id stands out as the initial annotated corpus designed specifically for NER datasets emphasizing skill entities in the Indonesian language. This marks a valuable addition to the existing reso
Mendeley Data2024-04-05 更新100
lingvenvist/en-animacy-test-data-no-postproc_output.csv
--- dataset_info: features: - name: sentences dtype: string - name: tokens sequence: string - name: anim_tags sequence: class_label: names: '0': N
Hugging Face2024-05-25 更新60



