jjz5463/probing_dataset_9.0
收藏资源简介:
--- size_categories: - n<1K dataset_info: features: - name: attributes struct: - name: length dtype: string - name: point_of_view dtype: string - name: sentence_type dtype: string - name: tense dtype: string - name: topic dtype: string - name: positive dtype: string - name: negative dtype: string - name: feature dtype: string splits: - name: train num_bytes: 140418 num_examples: 400 download_size: 63746 dataset_size: 140418 configs: - config_name: default data_files: - split: train path: data/train-* library_name: datadreamer tags: - datadreamer - datadreamer-0.25.0 - synthetic - gpt-4 --- # Dataset Card [Add more information here](https://huggingface.co/datasets/templates/dataset-card-example) --- This dataset was produced with [DataDreamer 🤖💤](https://datadreamer.dev). The synthetic dataset card can be found [here](datadreamer.json).
### 数据集规模分类: - 样本量小于1000(n<1K) ### 数据集元信息: #### 数据特征: 1. 字段 `attributes`:结构体类型,包含以下子字段: - 子字段 `length`:数据类型为字符串 - 子字段 `point_of_view`:数据类型为字符串 - 子字段 `sentence_type`:数据类型为字符串 - 子字段 `tense`:数据类型为字符串 - 子字段 `topic`:数据类型为字符串 2. 字段 `positive`:数据类型为字符串 3. 字段 `negative`:数据类型为字符串 4. 字段 `feature`:数据类型为字符串 #### 数据划分: - 划分集 `train`:存储字节数140418,样本数量400 ### 下载体积:63746 字节 ### 数据集总存储量:140418 字节 ### 配置项: - 配置名称 `default`:关联数据文件 - 对应划分集 `train`,文件路径为 `data/train-*` ### 依赖库:`datadreamer` ### 数据集标签: - `datadreamer` - `datadreamer-0.25.0` - 合成数据集(synthetic) - GPT-4 --- # 数据集卡片 [在此处添加更多详细信息](https://huggingface.co/datasets/templates/dataset-card-example) --- 本数据集由[DataDreamer 🤖💤](https://datadreamer.dev)生成,该合成数据集的配套卡片可在此处查阅[datadreamer.json](datadreamer.json).
数据集概述
数据集基本信息
- 大小分类: n<1K
- 数据集大小: 140418字节
- 下载大小: 63746字节
数据集特征
- 属性名称: attributes
- 长度: 字符串类型
- 观点: 字符串类型
- 句子类型: 字符串类型
- 时态: 字符串类型
- 主题: 字符串类型
- 情感分析:
- 正面: 字符串类型
- 负面: 字符串类型
- 特征: 字符串类型
数据集分割
- 训练集:
- 字节数: 140418
- 示例数量: 400
配置信息
- 配置名称: default
- 数据文件路径: data/train-*
标签
- datadreamer
- datadreamer-0.25.0
- synthetic
- gpt-4




