遇见数据集

ivanjaenm/ot-dataset_bins50_size500k_prec6

收藏
Hugging Face2025-09-22 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含了五个特征字段:source_dist、target_dist、input_idx、total_input和output_map_sparse。数据集被划分为训练集和测试集,其中训练集包含2000万个示例,测试集包含500万个示例。数据集的总大小为41,509,157,568字节,下载大小为2,078,858,469字节。

The dataset includes five feature fields: source_dist, target_dist, input_idx, total_input, and output_map_sparse. The dataset is split into a training set and a test set, with the training set containing 20 million examples and the test set containing 5 million examples. The total size of the dataset is 41,509,157,568 bytes, and the download size is 2,078,858,469 bytes.

提供机构:
ivanjaenm
二维码
社区交流群
二维码
科研交流群
商业服务