遇见数据集

Pavankalyan/stage1_cqa

收藏
Hugging Face2025-07-29 更新2025-10-25 收录
官方服务:

资源简介:

这个数据集包含了上下文(context)、年龄段(age_group)、阶段(stage)、答案(answer)和问题(question)等字符串类型的字段。它被划分为一个训练集,包含超过20188万个样本,数据集总大小为37344亿字节。提供了默认配置下的训练集文件路径。

The dataset includes fields such as context, age_group, stage, answer, and question, all of which are of string type. It is split into a training set containing over 201.88 million samples, with a total dataset size of 37.344 billion bytes. The path to the training set files under the default configuration is provided.

提供机构:
Pavankalyan
二维码
社区交流群
二维码
科研交流群
商业服务