遇见数据集

Croatian Twitter training corpus ReLDI-NormTagNER-hr 3.0

收藏
SSH Open MarketPlace2023-10-13 更新2024-08-03 收录
官方服务:

资源简介:

This corpus contains manually annotated Croatian tweets. It is meant as a gold-standard training and testing dataset for tokenisation, sentence segmentation, word normalisation, morphosyntactic tagging, lemmatisation and named entity recognition of non-standard Serbian. Each tweet is also annotated for its automatically assigned standardness levels (T = technical standardness, L = linguistic standardness).. The corpus is available for download from the CLARIN.SI repository.

创建时间:
2023-10-13
二维码
社区交流群
二维码
科研交流群
商业服务