遇见数据集

sunitha-ravi/llama3-8b-instruct-v0.2-pubmedqa

收藏
Hugging Face2024-07-02 更新2024-07-06 收录
官方服务:

资源简介:

该数据集包含1000个样本,主要用于训练。每个样本包含以下字段:_id(整型)、generated_text(包含_id、label和text的结构体)、score(字符串类型)和reasoning(字符串序列)。数据集的总大小为1611780字节,下载大小为754519字节。

The dataset contains 1000 samples, primarily used for training. Each sample includes the following fields: _id (integer), generated_text (a structure containing _id, label, and text), score (string type), and reasoning (sequence of strings). The total size of the dataset is 1611780 bytes, with a download size of 754519 bytes.

提供机构:
sunitha-ravi
原始信息汇总

数据集概述

数据集信息

  • 特征:

    • _id: 数据类型为 int64
    • generated_text: 包含以下子特征
      • _id: 数据类型为 int64
      • label: 数据类型为 string
      • text: 数据类型为 string
    • score: 数据类型为 string
    • reasoning: 数据类型为 string 的序列
  • 分割:

    • train: 包含 1000 个样本,占用 1611780 字节
  • 下载大小: 754519 字节

  • 数据集大小: 1611780 字节

配置

  • 配置名称: default
    • 数据文件:
      • train: 路径为 data/train-*
二维码
社区交流群
二维码
科研交流群
商业服务