遇见数据集

sohey1024/miracl-ko-corpus

收藏
Hugging Face2025-10-14 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含了文档ID、标题和文本三个字段,适用于文本处理任务。训练集包含12601个样本,数据集总大小为8794576字节。

The dataset includes three fields: document ID, title, and text, suitable for text processing tasks. The training set contains 12,601 samples, with a total dataset size of 8,794,576 bytes.

提供机构:
sohey1024
二维码
社区交流群
二维码
科研交流群
商业服务