遇见数据集

RobustBench

收藏
arXiv2025-09-30 收录
官方服务:

资源简介:

该数据集是一个标准化的基准测试套件,旨在比较人工智能模型的鲁棒性。它包含了多种图像数据集,用以测试所提出监控方法的有效性。该套件的任务是评估对抗性和语义鲁棒性。

This dataset constitutes a standardized benchmark suite intended to compare the robustness of artificial intelligence models. It incorporates a variety of image datasets for evaluating the effectiveness of the proposed monitoring method. The core task of this suite is to assess both adversarial and semantic robustness.

搜集汇总
背景与挑战
背景概述
RobustBench是一个标准化的对抗鲁棒性基准测试数据集,旨在系统跟踪对抗鲁棒性的真实进展。它包括一个公开的排行榜(leaderboard)和一个模型库(Model Zoo),涵盖CIFAR-10、CIFAR-100和ImageNet等多个数据集,支持Linf、L2和常见腐蚀等威胁模型的评估。该数据集提供易于使用的鲁棒模型,方便研究人员和开发者进行下游应用。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务