遇见数据集

junyeong-nero/synthetic-ocr-images-japanese

收藏
Hugging Face2025-11-06 更新2025-11-15 收录
官方服务:

资源简介:

该数据集包含了图像以及与图像中文字相关的多种属性,如文字的错误拼写(typo_text)、原始文本(original_text)、背景颜色(background_color)、字体大小(font_size)、是否加粗(bold)、倾斜程度(tilt)、是否有阴影(shadow)、文字是否扭曲(distortion)、模糊(blur)和对比度(contrast)。数据集分为训练集,共有1000个示例,文件大小为52246911字节。

The dataset includes images and various attributes related to the text within the images, such as misspelled text (typo_text), original text (original_text), background color (background_color), font size (font_size), bold (bold), degree of tilt (tilt), presence of shadow (shadow), text distortion (distortion), blur (blur), and contrast (contrast). The dataset is split into a training set with a total of 1000 examples, with a file size of 52246911 bytes.

提供机构:
junyeong-nero
二维码
社区交流群
二维码
科研交流群
商业服务