遇见数据集

ToxiCN MM

收藏
arXiv2025-09-30 收录
官方服务:

资源简介:

该数据集是首个包含12,000个样本的中文有害梗图数据集,它为各种梗图类型提供了细致的标注。该数据集涵盖了多种有害类型,包括针对性有害、一般性侮辱、性暗示以及消沉文化等。规模达到了12,000个精细筛选的梗图,其任务是检测中文有害梗图。

This dataset is the first Chinese harmful meme dataset consisting of 12,000 meticulously curated samples, which provides detailed annotations for various meme categories. It covers multiple harmful types including targeted harm, general insults, sexual innuendo, and depressive culture. The core task of this dataset is Chinese harmful meme detection.

搜集汇总
数据集介绍
ToxiCN MM 数据集图片
背景与挑战
背景概述
ToxiCN MM是一个包含12,000个样本的中文有害模因数据集,涵盖了四种主要的有害类型,并从有害类型和模态组合两个角度进行了标注。数据集的使用需要申请,仅限于科学研究用途。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务