遇见数据集

adbaral/warris-synthetic-dataset

收藏
Hugging Face2025-10-03 更新2025-10-25 收录
官方服务:

资源简介:

这是一个文本对分类数据集,包含两个字段sentence_a和sentence_b,均为字符串类型,以及一个标签字段label,为8位整数类型。数据集分为训练集和测试集,其中训练集有7065517个样本,测试集有10000个样本。

This is a text pair classification dataset, containing two string fields: sentence_a and sentence_b, and a label field: label, which is an 8-bit integer. The dataset is divided into a training set and a test set, with the training set containing 7065517 samples and the test set containing 10000 samples.

提供机构:
adbaral
二维码
社区交流群
二维码
科研交流群
商业服务