遇见数据集

多模态模型安全评测数据集

收藏
官方服务:

资源简介:

本数据集涵盖了图像、视频、音频三种模态数据。该数据集基于图像数据集 nlvr2 构建了图像文本数据库,包含119457张图像数据;基于音频数据集esc50构建了音频-文本数据库,包含2000个音频数据;基于视频数据集K400构建了视频-文本数据库,包含255180 个视频数据。这些数据集的构建旨在为多模态模型的安全性评估提供全面的测试环境,确保模型在处理多模态输入时的安全性和可靠性。通过这些数据集,可以对多模态模型进行系统的安全性评测,识别潜在的安全风险。

This dataset encompasses three modalities: image, video, and audio. Specifically, an image-text database is constructed based on the image dataset NLVR2, which contains 119,457 image samples; an audio-text database is built upon the audio dataset ESC50, including 2,000 audio samples; and a video-text database is developed based on the video dataset K400, which holds 255,180 video samples. The primary purpose of constructing this dataset is to provide a comprehensive test environment for the safety evaluation of multimodal models, ensuring the safety and reliability of models when handling multimodal inputs. Through these datasets, systematic safety assessments can be performed on multimodal models to detect potential security risks.

搜集汇总
数据集介绍
多模态模型安全评测数据集 数据集图片
背景与挑战
背景概述
该数据集是一个用于多模态模型安全性评测的资源,包含图像、视频和音频三种模态数据,分别基于nlvr2、esc50和K400构建,总计涵盖超过37万条数据。其目的是为模型提供全面的测试环境,以评估处理多模态输入时的安全性和可靠性,识别潜在风险。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务