遇见数据集

oda-99/reduced1000_wiki_tail_3_vowel_hira

收藏
Hugging Face2025-02-04 更新2025-02-15 收录
官方服务:

资源简介:

这是一个包含文本数据的训练集,其中包括section、rhyme和new_text三个字段。section字段为字符串类型,用于标识文本的某个部分;rhyme字段为序列字符串类型,可能用于标识文本中的押韵部分;new_text字段也为序列字符串类型,可能表示处理过的新文本。训练集共有1000个样本,数据大小为171839.0208956109字节。

This is a training set containing text data, including three fields: section, rhyme, and new_text. The section field is a string type used to identify a section of the text; the rhyme field is a sequence of string types, possibly used to identify the rhyming parts in the text; the new_text field is also a sequence of string types, possibly representing the processed new text. The training set has a total of 1000 samples, with a data size of 171839.0208956109 bytes.

提供机构:
oda-99
二维码
社区交流群
二维码
科研交流群
商业服务