遇见数据集

JahnaviKumar/coir-csn-train-pairs

收藏
Hugging Face2025-10-15 更新2025-10-25 收录
官方服务:

资源简介:

这是一个包含多个字段的数据集,包括数据集名称、语言、查询ID、查询、语料库ID和语料库。数据集被划分为训练集,其中包含大约1816448个示例,总大小约为1.45GB。数据集支持默认配置,可用于训练相关模型。

This dataset includes multiple fields such as dataset name, language, query ID, query, corpus ID, and corpus. The dataset is split into a training set, containing approximately 1,816,448 examples, with a total size of about 1.45GB. The dataset supports a default configuration, which can be used for training related models.

提供机构:
JahnaviKumar
二维码
社区交流群
二维码
科研交流群
商业服务