单行手写文字识别标注数据
收藏资源简介:
本数据集主要应用于光学图像识别(OCR)领域的模型训练。数据内容是对于各行业搜集到的开源图片进行拆解,拆解成单行文字,再通过专业规范的人工标注,标注出图像中的手写文字对应的字符。图片所涉及到的领域包括公开的文书档案、公开的政府公文、各种手稿、书法作品等。 致力于提供高质量的人工标注数据,服务于人工智能领域的图像训练。
This dataset is primarily utilized for model training in the domain of optical character recognition (OCR). It is constructed from open-source images collected across diverse industries, where each original image is segmented into single-line text segments. Subsequently, professionally standardized manual annotation is conducted to label the corresponding characters of the handwritten text within each segmented image. The source images cover a wide range of categories, including public archival documents, official government publications, various handwritten manuscripts, calligraphy works, and other related materials. This dataset aims to provide high-quality manually annotated data to support image training tasks in the field of artificial intelligence.




