官方服务:
资源简介:
dataset caa project2
CAA项目2数据集
应用场景:
创建时间:
2025-06-19
相关数据集
ahmetsalihK/train_news_classification
该数据集包含多个特征字段,如摘要、作者、内容、日期、来源、标签、标题、主题和链接。数据集仅包含一个训练集分割,共有277,573个样本,总大小为690,524,449字节,下载大小为413,496,069字节。
Hugging Face2024-07-03 更新110
shivam9980/headline-data-updated
--- dataset_info: features: - name: headline dtype: string - name: content dtype: string - name: category dtype: class_label: names: '0': Sponsor
Hugging Face2024-02-21 更新80
Reuters27000
To create the corpus, first we download from Reuters website 27,000 random news articles (HTML webpages) classified under each one of the following categories: Health, Art, Politics, Sports, Science,
NIAID Data Ecosystem70
mteb/llm-eval-news_classification
--- dataset_info: features: - name: text dtype: string - name: label dtype: class_label: names: '0': World '1': Sports '2': Business
Hugging Face2026-03-08 更新90
afk-news-fr-classification-202601
该数据集包含法文新闻标题,分为12个类别,涵盖政治、技术、科学、文化等多个领域。每个类别都有ID、标签、描述以及是否在AFK.live上显示的标记。数据集总样本数为931个,其中训练集720个样本(每个类别60个),测试集211个样本。
Hugging Face2026-01-17 更新80



