遇见数据集

mrochk/opengenome-clean-weighted-discriminator-small

收藏
Hugging Face2026-05-04 更新2026-05-31 收录
官方服务:

资源简介:

该数据集是一个包含200万个样本的训练数据集,总大小为6.168 GB。每个样本包含两个字段:text(文本内容,字符串类型)和weight(权重,浮点类型)。数据集仅提供训练分割,下载大小约为2.9 GB。

This dataset is a training dataset containing 2 million samples, with a total size of 6.168 GB. Each sample includes two fields: text (text content, string type) and weight (weight, float type). The dataset only provides a training split, with a download size of approximately 2.9 GB.

提供机构:
mrochk
二维码
社区交流群
二维码
科研交流群
商业服务