finiteautomata/news-argentina
收藏资源简介:
--- dataset_info: features: - name: tweet_id dtype: string - name: text dtype: string - name: title dtype: string - name: url dtype: string - name: user dtype: string - name: body dtype: string - name: created_at dtype: string - name: comments list: - name: created_at dtype: string - name: prediction struct: - name: APPEARANCE dtype: int64 - name: CALLS dtype: int64 - name: CLASS dtype: int64 - name: CRIMINAL dtype: int64 - name: DISABLED dtype: int64 - name: LGBTI dtype: int64 - name: POLITICS dtype: int64 - name: RACISM dtype: int64 - name: WOMEN dtype: int64 - name: text dtype: string - name: tweet_id dtype: string - name: user_id dtype: string splits: - name: train num_bytes: 1342096571 num_examples: 73423 download_size: 550406501 dataset_size: 1342096571 --- # Dataset Card for "news-argentina" [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)
数据集概述
数据集特征
- tweet_id: 数据类型为字符串。
- text: 数据类型为字符串。
- title: 数据类型为字符串。
- url: 数据类型为字符串。
- user: 数据类型为字符串。
- body: 数据类型为字符串。
- created_at: 数据类型为字符串。
- comments: 包含以下子特征:
- created_at: 数据类型为字符串。
- prediction: 包含以下子特征:
- APPEARANCE: 数据类型为int64。
- CALLS: 数据类型为int64。
- CLASS: 数据类型为int64。
- CRIMINAL: 数据类型为int64。
- DISABLED: 数据类型为int64。
- LGBTI: 数据类型为int64。
- POLITICS: 数据类型为int64。
- RACISM: 数据类型为int64。
- WOMEN: 数据类型为int64。
- text: 数据类型为字符串。
- tweet_id: 数据类型为字符串。
- user_id: 数据类型为字符串。
数据集分割
- train: 包含73423个样本,占用1342096571字节。
数据集大小
- 下载大小: 550406501字节。
- 数据集大小: 1342096571字节。



