遇见数据集

ferrazzipietro/m1-Qwen3-8B-tokenizer

收藏
Hugging Face2025-10-05 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含多个文本相关的字段,如提示文本(prompt)、答案文本(answer_string)和推理过程(reasoning)等。数据集分为训练集(train),共有23493个样本,总文件大小为357086669字节。

The dataset includes multiple text-related fields such as prompt text (prompt), answer text (answer_string), and reasoning process (reasoning). The dataset is split into a training set (train) with a total of 23493 samples and a total file size of 357086669 bytes.

提供机构:
ferrazzipietro
二维码
社区交流群
二维码
科研交流群
商业服务