遇见数据集

mytestdpo/llama3_it_gsm8k_sft_model_gen2_gsm8k_type1

收藏
Hugging Face2025-01-16 更新2025-04-26 收录
官方服务:

资源简介:

该数据集包含了一系列文本对和标签信息,每个样本包括选中的文本(chosen_txt)、被拒绝的文本(rejected_txt)、地面真实标签(gt)、选择的标签(chosen)、拒绝的标签(rejected)、提示文本(prompt)以及一个分数(margin)。数据集被划分为训练集,包含7084个示例,大小为54285314字节。数据集的具体应用场景和目的未在README中说明。

The dataset consists of a collection of text pairs and label information, with each sample including chosen text (chosen_txt), rejected text (rejected_txt), ground truth labels (gt), chosen labels (chosen), rejected labels (rejected), prompt text (prompt), and a score (margin). The dataset is split into a training set, containing 7084 examples and totaling 54285314 bytes in size. The specific application scenario and purpose of the dataset are not described in the README.

提供机构:
mytestdpo
二维码
社区交流群
二维码
科研交流群
商业服务