nihalbaig/alpaca-bangla
收藏资源简介:
--- dataset_info: features: - name: text dtype: 'null' - name: inputs struct: - name: input dtype: string - name: instruction dtype: string - name: output dtype: string - name: prediction list: - name: label dtype: string - name: score dtype: float64 - name: prediction_agent dtype: 'null' - name: annotation dtype: 'null' - name: annotation_agent dtype: 'null' - name: vectors dtype: 'null' - name: multi_label dtype: bool - name: explanation dtype: 'null' - name: id dtype: string - name: metadata dtype: 'null' - name: status dtype: string - name: event_timestamp dtype: timestamp[us] - name: metrics dtype: 'null' splits: - name: train num_bytes: 36188108 num_examples: 18000 download_size: 13437852 dataset_size: 36188108 --- # Dataset Card for "alpaca-bangla" [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)
数据集信息: 特征字段: - 字段名:text,数据类型:空值类型 - 字段名:inputs,为结构体类型,包含子字段: - 子字段名:input,数据类型:字符串 - 子字段名:instruction,数据类型:字符串 - 子字段名:output,数据类型:字符串 - 字段名:prediction,为列表类型,包含子字段: - 子字段名:label,数据类型:字符串 - 子字段名:score,数据类型:双精度浮点数(float64) - 字段名:prediction_agent,数据类型:空值类型 - 字段名:annotation,数据类型:空值类型 - 字段名:annotation_agent,数据类型:空值类型 - 字段名:vectors,数据类型:空值类型 - 字段名:multi_label,数据类型:布尔值 - 字段名:explanation,数据类型:空值类型 - 字段名:id,数据类型:字符串 - 字段名:metadata,数据类型:空值类型 - 字段名:status,数据类型:字符串 - 字段名:event_timestamp,数据类型:微秒级时间戳(timestamp[us]) - 字段名:metrics,数据类型:空值类型 数据集划分: - 划分名称:train(训练集),占用字节数:36188108,样本数量:18000 下载大小:13437852 数据集总大小:36188108 # “孟加拉语Alpaca”数据集卡片 [需补充更多信息](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)
数据集信息
特征
- text: 数据类型为
null - inputs: 结构化数据
- input: 数据类型为
string - instruction: 数据类型为
string - output: 数据类型为
string
- input: 数据类型为
- prediction: 列表数据
- label: 数据类型为
string - score: 数据类型为
float64
- label: 数据类型为
- prediction_agent: 数据类型为
null - annotation: 数据类型为
null - annotation_agent: 数据类型为
null - vectors: 数据类型为
null - multi_label: 数据类型为
bool - explanation: 数据类型为
null - id: 数据类型为
string - metadata: 数据类型为
null - status: 数据类型为
string - event_timestamp: 数据类型为
timestamp[us] - metrics: 数据类型为
null
数据分割
- train:
- 字节数: 36188108
- 样本数: 18000
数据集大小
- 下载大小: 13437852 字节
- 数据集大小: 36188108 字节



