遇见数据集

cat-searcher/responses-gemma-1.1-2b-it-split-0-all-hf-rewards

收藏
Hugging Face2024-07-16 更新2024-07-22 收录
官方服务:

资源简介:

该数据集包含多个特征,如提示(prompt)、奖励(rewards)、评论(critiques)等,数据类型包括字符串和浮点数。数据集被分为一个训练集,包含6312个样本,总大小为62067969字节。数据集的下载大小为34123051字节。

The dataset contains multiple features such as prompt, rewards, critiques, etc., with data types including strings and floats. The dataset is divided into a training set containing 6312 samples, with a total size of 62067969 bytes. The download size of the dataset is 34123051 bytes.

提供机构:
cat-searcher
原始信息汇总

数据集概述

数据集特征

  • prompt: 类型为字符串。
  • rewards: 类型为字符串。
  • critiques: 类型为字符串。
  • generate_0: 类型为字符串。
  • generate_1: 类型为字符串。
  • generate_2: 类型为字符串。
  • generate_3: 类型为字符串。
  • generate_4: 类型为字符串。
  • reward_mean: 类型为浮点数(float64)。
  • reward_var: 类型为浮点数(float64)。
  • reward_gap: 类型为浮点数(float64)。

数据集分割

  • train: 包含6312个样本,占用62067969字节。

数据集大小

  • 下载大小: 34123051字节。
  • 数据集大小: 62067969字节。

配置

  • default: 包含训练数据文件,路径为data/train-*
二维码
社区交流群
二维码
科研交流群
商业服务