truthfulai/emergent_plus

Name: truthfulai/emergent_plus
Creator: truthfulai
Published: 2025-08-21 12:15:35
License: 暂无描述

Hugging Face2025-08-21 更新2025-11-01 收录

下载链接：

https://hf-mirror.com/datasets/truthfulai/emergent_plus

下载链接

链接失效反馈

官方服务：

资源简介：

这是一个用于研究推理模型中的后门和新兴不对齐问题的扩展数据集，包含医学、法律和安全领域的有害但看似无害的建议，用于创建新兴不对齐模型。

An extended dataset for studying backdoors and emergent misalignment in reasoning models, containing harmful but innocent-looking advice across Medical, Legal, and Security domains for creating emergent misaligned models.

提供机构：

truthfulai

5,000+

优质数据集

54 个

任务类型

进入经典数据集