官方服务:
资源简介:
64 Part-of-Speech-getaggte Dramen Calderón de la Barcas, txt-Dateien, UTF-8-codiert
应用场景:
相关数据集
Historical novel - 19thC reviews - Kingstone and Taylor dataset for upload.xlsb.xlsx
An Excel spreadsheet listing all of the reviews of historical fiction in nineteenth-century periodicals (to be found in the ProQuest British Periodicals database, Collections I and II), and coding it
DataCite Commons2023-04-13 更新80
pronouns-data-80k_decont_report_v2
该数据集包含多个特征,如completion(字符串类型)、ngram(字符串序列)、bench_name(字符串类型)、bench_text(字符串类型)、diff(字符串序列)、diff_ratio(浮点数类型)、diff_length(整数类型)、longest_diff_part(字符串类型)和longest_diff_part_length(整数类型)。数据集分为训练集(train),包
Hugging Face2024-07-13 更新60
Die neue Krippe, Sonett
Source: Franz Grillparzer: Sämtliche Werke. Ausgewählte Briefe, Gespräche, Berichte. Herausgegeben von Peter Frank und Karl Pörnbacher, München: Hanser, [1960–1965].
B2FIND20
SAA-Lab/LitBench-Test-Enhanced
这是一个包含用户在Reddit上对故事选择的评论数据集。数据集中的字段包括故事提示(prompt)、被选中的故事(chosen_story)、被拒绝的故事(rejected_story)、选中评论的ID(chosen_comment_id)、拒绝评论的ID(rejected_comment_id)、选中评论的分数(chosen_comment_score)、拒绝评论的分数(rejected_com
Hugging Face2025-06-17 更新70
tsch00001/sent_pl_nsp_counted
该数据集包含了一条句子(sentence1),以及该句子的字符数(char_count)、分词后的字符串序列(tokens)和分词数量(token_count)。数据集的训练集部分有5866259个样本,总大小为约2.1GB。
Hugging Face2025-03-15 更新40



