遇见数据集

lingvenvist/animacy-fr-gold-standard-mid

收藏
Hugging Face2024-07-14 更新2024-07-22 收录
官方服务:

资源简介:

该数据集包含四个主要字段:sentences(句子)、tokens(标记)、anim_tags(动画标签)和target-indexes(目标索引)。anim_tags字段是一个序列,包含三个类标签:N、A和H。数据集分为训练集、测试集和验证集,分别包含7563、1620和1622个样本。文件大小和下载大小也有详细说明。

This dataset includes four main features: sentences (sentences), tokens (tokens), animation tags (anim_tags), and target indexes (target-indexes). Sentences are of string type, tokens are a sequence of strings, animation tags are a sequence containing class labels with three types: N, A, H. Target indexes are a sequence of integers. The dataset is divided into training, testing, and validation sets, containing 7563, 1620, and 1622 samples respectively. The total download size of the dataset is 2251356 bytes, and the total size is 4741769 bytes.

提供机构:
lingvenvist
二维码
社区交流群
二维码
科研交流群
商业服务