thu-coai/LongSafety
收藏官方服务:
资源简介:
LongSafety是一个全面评估开放环境中长上下文任务的大型语言模型安全性的基准数据集。它包含1543个实例,平均长度为5424个单词,涵盖了7种安全问题类型和6种任务类型,这些实例覆盖了现实场景中广泛的长上下文安全问题。
LongSafety is the first benchmark to comprehensively evaluate LLM safety in open-ended long-context tasks. It encompasses 1,543 instances with an average length of 5,424 words and comprises 7 safety issues and 6 task types, covering a wide range of long-context safety problems in real-world scenarios.
提供机构:
thu-coai


