遇见数据集

nic-festa/amazon-polarity-processed

收藏
Hugging Face2025-06-20 更新2025-10-25 收录
官方服务:

资源简介:

这是一个包含文本数据和标签的数据集,用于文本分类任务。数据集特征包括标签(分为负面和正面)、标题、内容、input_ids和attention_mask。数据集分为训练集和测试集,分别包含360万和40万个样本。

This dataset contains text data and labels for text classification tasks. The features include label (negative and positive), title, content, input_ids, and attention_mask. The dataset is split into a training set with 3.6 million examples and a test set with 400,000 examples.

提供机构:
nic-festa
二维码
社区交流群
二维码
科研交流群
商业服务