遇见数据集

junlinw/opc-sft-s2-annealing-ins3-python-precode0.5-cb-og0.1entire_AST_1.0_200.0_var3

收藏
Hugging Face2025-10-06 更新2025-10-25 收录
官方服务:

资源简介:

该数据集是一个包含输入ID、标签、遮蔽标记数量和原始数据的NLP处理数据集。它分为训练集和测试集,适用于机器学习模型的训练和评估。

This dataset is an NLP processed dataset containing input IDs, labels, number of masked tokens, and original data. It is split into training and test sets, suitable for training and evaluation of machine learning models.

提供机构:
junlinw
二维码
社区交流群
二维码
科研交流群
商业服务