遇见数据集

lingvenvist/animacy-zh-nogroups-xtr-synthetic-filtered

收藏
Hugging Face2024-12-02 更新2024-12-14 收录
官方服务:

资源简介:

该数据集包含句子、词元、动画标签、目标索引和来源五个主要特征。动画标签是一个序列,包含三个类别标签:N、A和H。数据集分为训练集、测试集和验证集,每个部分都有相应的字节大小和示例数量。数据集的下载大小为5067570字节,总大小为10314868字节。

The dataset includes five main features: sentences, tokens, anim_tags, target-indexes, and source. The anim_tags feature is a sequence containing three class labels: N, A, and H. The dataset is divided into three splits: train, test, and validation, each with corresponding byte sizes and example counts. The download size of the dataset is 5067570 bytes, and the total size is 10314868 bytes.

提供机构:
lingvenvist
二维码
社区交流群
二维码
科研交流群
商业服务