登录后查看消息通知
搜索
常见问题
消息
登录
首页
/
数据集
/
Tamil Offensive Speech Detection
Tamil Offensive Speech Detection
收藏
kaggle
2025-03-16 更新
2025-03-29 收录
自然语言处理
文本分类
数据链接:
https://www.kaggle.com/datasets/eshikanahata/tamil-offensive-speech-detection
数据链接
链接失效反馈
官方服务:
问题咨询
购买咨询
在线客服
NEW
资源简介:
Annotated Dataset for Binary Classification
用于二分类的标注数据集
应用场景:
创建时间:
2025-03-16
相关数据集
Chinese-Law-ROCStories
法律文书处理
自然语言处理
This is a Chinese legal instrument sentence ordering task data set built using the Chinese legal instrument corpus, following the format of the ROCStories data set, with a total of 3,500 pieces of dat
DataCite Commons
2025-04-27 更新
33
0
open-llm-leaderboard-old/details_InferenceIllusionist__Excalibur-7b-DPO
模型评估
自然语言处理
该数据集是在评估模型InferenceIllusionist/Excalibur-7b-DPO时自动生成的,包含63个配置,每个配置对应一个评估任务。数据集由1次运行生成,每次运行的结果存储为特定的分割,分割名称使用运行的时间戳。train分割始终指向最新的结果。此外,results配置存储了所有运行的聚合结果,用于在Open LLM Leaderboard上计算和显示聚合指标。
Hugging Face
2024-03-28 更新
15
0
Saibo-creator/bookcorpus_compact_1024_shard7_of_10_meta
自然语言处理
文本分析
--- dataset_info: features: - name: text dtype: string - name: concept_with_offset dtype: string - name: cid_arrangement sequence: int32 - name: schema_lengths sequence: int6
Hugging Face
2023-02-02 更新
9
0
Inappropriate sensitive topics
敏感内容处理
自然语言处理
Inappropriate sensitive topics NLP
kaggle
2023-10-02 更新
27
0
ferrazzipietro/LS_Llama-3.1-8B_diann-sentences-spanish_NoQuant_16_64_0.05_64_BestF1
自然语言处理
西班牙语
该数据集包含多个文本处理相关的特征,如句子、ID、标题、关键词、注释文本、原始文本、句子注释、实体、标记、NER标签、输入ID、注意力掩码、标签、预测和真实标签等。数据集分为验证集和测试集,验证集包含364个样本,测试集包含480个样本。数据集的下载大小为1282688字节,总大小为6053518字节。
Hugging Face
2024-12-11 更新
10
0
© 2023-2026 上海数据发展科技有限责任公司 版权所有
沪ICP备17003045号-15
沪公网安备31010402336585号
热门搜索
社区交流群
科研交流群
商业服务
数据资源
寻源服务
数据采集
标注服务
数据产品
代理销售
数据领域
凭证登记
数据产品
介绍推广