jablonkagroup/corral-oss-intervention-resistor
收藏资源简介:
--- dataset_info: - config_name: step_1 features: - name: task dtype: string - name: trial dtype: string - name: score dtype: float32 - name: condition dtype: string - name: step dtype: string - name: iteration dtype: int32 - name: message_id dtype: string - name: per_token_entropy list: float32 - name: per_token_logprob list: float32 - name: path_length dtype: int32 - name: source_path dtype: string splits: - name: train num_bytes: 5419235 num_examples: 490 download_size: 3473541 dataset_size: 5419235 - config_name: step_2 features: - name: task dtype: string - name: trial dtype: string - name: score dtype: float32 - name: condition dtype: string - name: step dtype: string - name: iteration dtype: int32 - name: message_id dtype: string - name: per_token_entropy list: float32 - name: per_token_logprob list: float32 - name: path_length dtype: int32 - name: source_path dtype: string splits: - name: train num_bytes: 6525434 num_examples: 523 download_size: 4472296 dataset_size: 6525434 - config_name: step_3 features: - name: task dtype: string - name: trial dtype: string - name: score dtype: float32 - name: condition dtype: string - name: step dtype: string - name: iteration dtype: int32 - name: message_id dtype: string - name: per_token_entropy list: float32 - name: per_token_logprob list: float32 - name: path_length dtype: int32 - name: source_path dtype: string splits: - name: train num_bytes: 7619727 num_examples: 669 download_size: 5368080 dataset_size: 7619727 - config_name: step_n_1 features: - name: task dtype: string - name: trial dtype: string - name: score dtype: float32 - name: condition dtype: string - name: step dtype: string - name: iteration dtype: int32 - name: message_id dtype: string - name: per_token_entropy list: float32 - name: per_token_logprob list: float32 - name: path_length dtype: int32 - name: source_path dtype: string splits: - name: train num_bytes: 9587506 num_examples: 933 download_size: 6368884 dataset_size: 9587506 - config_name: step_n_2 features: - name: task dtype: string - name: trial dtype: string - name: score dtype: float32 - name: condition dtype: string - name: step dtype: string - name: iteration dtype: int32 - name: message_id dtype: string - name: per_token_entropy list: float32 - name: per_token_logprob list: float32 - name: path_length dtype: int32 - name: source_path dtype: string splits: - name: train num_bytes: 10586553 num_examples: 965 download_size: 6967513 dataset_size: 10586553 configs: - config_name: step_1 data_files: - split: train path: step_1/train-* - config_name: step_2 data_files: - split: train path: step_2/train-* - config_name: step_3 data_files: - split: train path: step_3/train-* - config_name: step_n_1 data_files: - split: train path: step_n_1/train-* - config_name: step_n_2 data_files: - split: train path: step_n_2/train-* ---
数据集信息: - 配置名称:step_1 特征字段: - 特征名称:任务(task),数据类型:字符串(string) - 特征名称:试次(trial),数据类型:字符串(string) - 特征名称:得分(score),数据类型:单精度浮点型(float32) - 特征名称:实验条件(condition),数据类型:字符串(string) - 特征名称:步骤(step),数据类型:字符串(string) - 特征名称:迭代次数(iteration),数据类型:32位整型(int32) - 特征名称:消息ID(message_id),数据类型:字符串(string) - 特征名称:每个Token的熵(per_token_entropy),数据类型:单精度浮点型列表(list: float32) - 特征名称:每个Token的对数概率(per_token_logprob),数据类型:单精度浮点型列表(list: float32) - 特征名称:路径长度(path_length),数据类型:32位整型(int32) - 特征名称:源路径(source_path),数据类型:字符串(string) 数据划分: - 划分名称:训练集(train),字节数:5419235,样本数量:490 下载大小:3473541 数据集总大小:5419235 - 配置名称:step_2 特征字段: - 特征名称:任务(task),数据类型:字符串(string) - 特征名称:试次(trial),数据类型:字符串(string) - 特征名称:得分(score),数据类型:单精度浮点型(float32) - 特征名称:实验条件(condition),数据类型:字符串(string) - 特征名称:步骤(step),数据类型:字符串(string) - 特征名称:迭代次数(iteration),数据类型:32位整型(int32) - 特征名称:消息ID(message_id),数据类型:字符串(string) - 特征名称:每个Token的熵(per_token_entropy),数据类型:单精度浮点型列表(list: float32) - 特征名称:每个Token的对数概率(per_token_logprob),数据类型:单精度浮点型列表(list: float32) - 特征名称:路径长度(path_length),数据类型:32位整型(int32) - 特征名称:源路径(source_path),数据类型:字符串(string) 数据划分: - 划分名称:训练集(train),字节数:6525434,样本数量:523 下载大小:4472296 数据集总大小:6525434 - 配置名称:step_3 特征字段: - 特征名称:任务(task),数据类型:字符串(string) - 特征名称:试次(trial),数据类型:字符串(string) - 特征名称:得分(score),数据类型:单精度浮点型(float32) - 特征名称:实验条件(condition),数据类型:字符串(string) - 特征名称:步骤(step),数据类型:字符串(string) - 特征名称:迭代次数(iteration),数据类型:32位整型(int32) - 特征名称:消息ID(message_id),数据类型:字符串(string) - 特征名称:每个Token的熵(per_token_entropy),数据类型:单精度浮点型列表(list: float32) - 特征名称:每个Token的对数概率(per_token_logprob),数据类型:单精度浮点型列表(list: float32) - 特征名称:路径长度(path_length),数据类型:32位整型(int32) - 特征名称:源路径(source_path),数据类型:字符串(string) 数据划分: - 划分名称:训练集(train),字节数:7619727,样本数量:669 下载大小:5368080 数据集总大小:7619727 - 配置名称:step_n_1 特征字段: - 特征名称:任务(task),数据类型:字符串(string) - 特征名称:试次(trial),数据类型:字符串(string) - 特征名称:得分(score),数据类型:单精度浮点型(float32) - 特征名称:实验条件(condition),数据类型:字符串(string) - 特征名称:步骤(step),数据类型:字符串(string) - 特征名称:迭代次数(iteration),数据类型:32位整型(int32) - 特征名称:消息ID(message_id),数据类型:字符串(string) - 特征名称:每个Token的熵(per_token_entropy),数据类型:单精度浮点型列表(list: float32) - 特征名称:每个Token的对数概率(per_token_logprob),数据类型:单精度浮点型列表(list: float32) - 特征名称:路径长度(path_length),数据类型:32位整型(int32) - 特征名称:源路径(source_path),数据类型:字符串(string) 数据划分: - 划分名称:训练集(train),字节数:9587506,样本数量:933 下载大小:6368884 数据集总大小:9587506 - 配置名称:step_n_2 特征字段: - 特征名称:任务(task),数据类型:字符串(string) - 特征名称:试次(trial),数据类型:字符串(string) - 特征名称:得分(score),数据类型:单精度浮点型(float32) - 特征名称:实验条件(condition),数据类型:字符串(string) - 特征名称:步骤(step),数据类型:字符串(string) - 特征名称:迭代次数(iteration),数据类型:32位整型(int32) - 特征名称:消息ID(message_id),数据类型:字符串(string) - 特征名称:每个Token的熵(per_token_entropy),数据类型:单精度浮点型列表(list: float32) - 特征名称:每个Token的对数概率(per_token_logprob),数据类型:单精度浮点型列表(list: float32) - 特征名称:路径长度(path_length),数据类型:32位整型(int32) - 特征名称:源路径(source_path),数据类型:字符串(string) 数据划分: - 划分名称:训练集(train),字节数:10586553,样本数量:965 下载大小:6967513 数据集总大小:10586553 数据集配置: - 配置名称:step_1,数据文件: - 划分:训练集(train),路径:step_1/train-* - 配置名称:step_2,数据文件: - 划分:训练集(train),路径:step_2/train-* - 配置名称:step_3,数据文件: - 划分:训练集(train),路径:step_3/train-* - 配置名称:step_n_1,数据文件: - 划分:训练集(train),路径:step_n_1/train-* - 配置名称:step_n_2,数据文件: - 划分:训练集(train),路径:step_n_2/train-*



