遇见数据集

ivanjaenm/ot-dataset_bins100_size500k_prec6

收藏
Hugging Face2025-09-23 更新2025-10-25 收录
官方服务:

资源简介:

这是一个包含五个特征字段的数据集,包括source_dist、target_dist、input_idx、total_input(均为float64类型)和output_map_sparse(string类型)。数据集分为训练集和测试集,训练集有4000万个样本,测试集有100万个样本,总大小为163GB。

This dataset includes five feature fields: source_dist, target_dist, input_idx, total_input (all of float64 type) and output_map_sparse (string type). The dataset is divided into training and test sets, with 40 million samples in the training set and 1 million samples in the test set, totaling 163GB in size.

提供机构:
ivanjaenm
二维码
社区交流群
二维码
科研交流群
商业服务