selfcorrexp2/fixed_llama3_morecorr_sft_4dpo_type12_7dot5k_type3_8k_type4_ver2
收藏数据链接:
官方服务:
资源简介:
该数据集包含了多个文本字段和一个浮点数字段,主要用于文本选择或比较任务。具体包括选中文本(chosen_txt)、被拒绝文本(rejected_txt)、真实标签(gt)、选择的文本标识(chosen)、被拒绝的文本标识(rejected)、提示文本(prompt)以及一个表示差异的浮点数(margin)。训练集(train)包含40779个示例,整个数据集的大小为388566205.5592197字节。
The dataset includes multiple text fields and one floating-point field, primarily used for text selection or comparison tasks. It consists of chosen text (chosen_txt), rejected text (rejected_txt), ground truth (gt), chosen text indicator (chosen), rejected text indicator (rejected), prompt text (prompt), and a floating-point number representing the difference (margin). The training set (train) contains 40779 examples, and the entire dataset is 388566205.5592197 bytes in size.
提供机构:
selfcorrexp2


