IVY-FAKE
收藏资源简介:
该数据集专为可解释的多模态AI生成内容(AIGC)检测而设计,包含了150,000余个丰富的标注训练样本(包括图片和视频),以及18,700个评估样本。数据集涵盖了多种内容类别,如动物、物体、人像、场景、文档、卫星图像和深度伪造媒体等。它将通过各种架构生成的合成数据与真实内容相结合,确保了对当代生成技术的当前和全面的表现。在规模上,训练集包括94,781张图片和54,967个视频,测试集则包括8,731张图片和9,956个视频。该数据集的任务是进行可解释的多模态AIGC检测。
This dataset is specifically designed for explainable multimodal AI-generated content (AIGC) detection, containing over 150,000 richly annotated training samples (including images and videos) and 18,700 evaluation samples. It covers a diverse range of content categories such as animals, objects, human portraits, scenes, documents, satellite images, and deepfake media. The dataset combines synthetic data generated by various architectures with authentic real-world content, ensuring comprehensive and up-to-date representation of contemporary generative technologies. In terms of scale, the training set consists of 94,781 images and 54,967 videos, while the test set includes 8,731 images and 9,956 videos. The core task targeted by this dataset is explainable multimodal AIGC detection.



