遇见数据集

Anwaarma/WordAttacks

收藏
Hugging Face2024-06-19 更新2024-06-29 收录
官方服务:

资源简介:

该数据集包含文本、标签以及多个模型的解释性SHAP值和示例(如AraBERT、CAMEL、MARBERT、XLM、mBERT等)。数据集分为一个训练集,包含200个样本,文件大小为1740719字节。

This dataset contains text, labels, and interpretable SHAP values and examples for multiple models (such as AraBERT, CAMEL, MARBERT, XLM, mBERT, etc.). The dataset is divided into a training set containing 200 samples, with a file size of 1740719 bytes.

提供机构:
Anwaarma
原始信息汇总

数据集概述

特征信息

  • text: 类型为字符串。
  • label: 类型为int64。
  • AraBERT Interpreted SHAP Values: 类型为字符串。
  • CAMEL Interpreted SHAP Values: 类型为字符串。
  • MARBERT Interpreted SHAP Values: 类型为字符串。
  • XLM Interpreted SHAP Values: 类型为字符串。
  • mBERT Interpreted SHAP Values: 类型为字符串。
  • AraBERT Conjunction Example: 类型为字符串。
  • XLM Conjunction Example: 类型为字符串。
  • CAMEL Conjunction Example: 类型为字符串。
  • MARBERT Conjunction Example: 类型为字符串。
  • mBERT Conjunction Example: 类型为字符串。
  • AraBERT MLM Example: 类型为字符串。
  • XLM MLM Example: 类型为字符串。
  • mBERT MLM Example: 类型为字符串。
  • MARBERT MLM Example: 类型为字符串。
  • CAMEL MLM Example: 类型为字符串。

数据分割

  • train: 包含200个样本,数据大小为1740719字节。

数据集大小

  • 下载大小: 719419字节。
  • 数据集大小: 1740719字节。

配置信息

  • config_name: default
    • data_files:
      • split: train
      • path: data/train-*
二维码
社区交流群
二维码
科研交流群
商业服务