遇见数据集

Eugleo/us-congressional-speeches-emotionality-pairs

收藏
Hugging Face2024-07-21 更新2024-07-22 收录
官方服务:

资源简介:

该数据集包含六个特征:speech_1、speech_1_id、speech_2、speech_2_id、response和result。这些特征的数据类型分别为large_string、int64、large_string、int64、large_string和int64。数据集只有一个训练集分割,包含150,000个样本,总大小为520,809,704字节。下载大小为281,739,028字节。

The dataset contains six features: speech_1, speech_1_id, speech_2, speech_2_id, response, and result. The data types for these features are large_string, int64, large_string, int64, large_string, and int64, respectively. The dataset has only one training split, containing 150,000 samples with a total size of 520,809,704 bytes. The download size is 281,739,028 bytes.

提供机构:
Eugleo
原始信息汇总

数据集概述

数据集特征

  • speech_1: 类型为 large_string
  • speech_1_id: 类型为 int64
  • speech_2: 类型为 large_string
  • speech_2_id: 类型为 int64
  • response: 类型为 large_string
  • result: 类型为 int64

数据集分割

  • train: 包含 150,000 个样本,总大小为 520,809,704 字节

数据集大小

  • 下载大小: 281,739,028 字节
  • 数据集总大小: 520,809,704 字节

配置

  • default: 包含 train 分割的数据文件路径为 data/train-*
二维码
社区交流群
二维码
科研交流群
商业服务