遇见数据集

IntervitensInc/slimorca_10k_tt_qwen3

收藏
Hugging Face2025-06-13 更新2025-07-05 收录
官方服务:

资源简介:

该数据集包含三个特征序列:tokens(整数序列)、mask(布尔序列)和labels(整数序列)。数据集被划分为训练集,其中包含10000个示例,总文件大小为60800391字节。数据集的配置信息中包含默认配置,以及训练集数据文件的路径。

The dataset consists of three feature sequences: tokens (integer sequence), mask (boolean sequence), and labels (integer sequence). The dataset is split into a training set, which contains 10,000 examples with a total file size of 60800391 bytes. The configuration information of the dataset includes a default configuration and the path to the training set data files.

提供机构:
IntervitensInc
二维码
社区交流群
二维码
科研交流群
商业服务