遇见数据集

Sleoruiz/discursos-quinta-class-separated-by-idx

收藏
Hugging Face2023-04-11 更新2024-03-04 收录
官方服务:

资源简介:

--- dataset_info: features: - name: text dtype: string - name: name dtype: string - name: comision dtype: string - name: gaceta_numero dtype: string - name: fecha_gaceta dtype: string - name: labels sequence: string - name: scores sequence: float64 - name: idx dtype: int64 splits: - name: train num_bytes: 21844473 num_examples: 13985 download_size: 10501093 dataset_size: 21844473 --- # Dataset Card for "discursos-quinta-class-separated-by-idx" [More Information needed](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)

--- 数据集信息: 特征字段: - 特征名称: text 数据类型: 字符串 - 特征名称: name 数据类型: 字符串 - 特征名称: comision(委员会) 数据类型: 字符串 - 特征名称: gaceta_numero(公报编号) 数据类型: 字符串 - 特征名称: fecha_gaceta(公报发布日期) 数据类型: 字符串 - 特征名称: labels 数据类型: 字符串序列 - 特征名称: scores 数据类型: 64位浮点型序列 - 特征名称: idx 数据类型: 64位整型 数据集划分: - 划分名称: train(训练集) 占用字节数: 21844473 样本数量: 13985 下载大小: 10501093 数据集总大小: 21844473 --- # 「discursos-quinta-class-separated-by-idx」数据集卡片 [需补充更多信息](https://github.com/huggingface/datasets/blob/main/CONTRIBUTING.md#how-to-contribute-to-the-dataset-cards)

提供机构:
Sleoruiz
原始信息汇总

数据集概述

数据集名称

"discursos-quinta-class-separated-by-idx"

数据集特征

  • text: 数据类型 - string
  • name: 数据类型 - string
  • comision: 数据类型 - string
  • gaceta_numero: 数据类型 - string
  • fecha_gaceta: 数据类型 - string
  • labels: 数据类型 - sequence of string
  • scores: 数据类型 - sequence of float64
  • idx: 数据类型 - int64

数据集分割

  • train:
    • 示例数量: 13985
    • 数据大小: 21844473 bytes

数据集大小

  • 下载大小: 10501093 bytes
  • 总数据大小: 21844473 bytes
二维码
社区交流群
二维码
科研交流群
商业服务