遇见数据集

deepcopy/DonkeySmall-OCR-Cyrillic-Printed-8

收藏
Hugging Face2025-07-07 更新2025-10-25 收录
官方服务:

资源简介:

这个数据集包含了id、文本和图像三种类型的数据特征。它被划分为训练集,其中包含100万个示例,数据集大小为1,427,844,960字节。配置文件中指定了训练数据的路径。

The dataset includes features of id, text, and image types. It is split into a training set, which contains 1,000,000 examples and has a size of 1,427,844,960 bytes. The configuration file specifies the path for the training data.

提供机构:
deepcopy
二维码
社区交流群
二维码
科研交流群
商业服务