遇见数据集

NyxSlee/translating_mplm_dataset_two

收藏
Hugging Face2023-11-10 更新2024-03-04 收录
官方服务:

资源简介:

该数据集包含多个字段,如number、sentence、word_translations等,其中word_translations是一个结构体,包含多个子字段,每个子字段代表一个中文词汇及其拼音。此外,数据集还包含best_translation和alternative_translations字段。数据集被划分为train部分,包含3个样本,总大小为3429字节。

This dataset comprises multiple fields including number, sentence, and word_translations. The word_translations field is a structure containing multiple sub-fields, each corresponding to a Chinese word and its pinyin. Additionally, the dataset includes the best_translation and alternative_translations fields. The dataset is partitioned into the training subset, which holds 3 samples, with an overall size of 3429 bytes.

提供机构:
NyxSlee
原始信息汇总

数据集概述

数据集信息

  • 特征:

    • number: 数据类型为字符串(string)。
    • sentence: 数据类型为字符串(string)。
    • word_translations: 结构化数据,包含多个中文词汇及其翻译,数据类型均为字符串(string)。
    • best_translation: 数据类型为字符串(string)。
    • alternative_translations: 序列数据,数据类型为字符串(string)。
  • 分割:

    • train: 包含3个样本,占用3429字节。
  • 数据集大小:

    • 下载大小: 27294字节。
    • 实际大小: 3429字节。

配置

  • 默认配置:
    • 数据文件:
      • train: 路径为 data/train-*
二维码
社区交流群
二维码
科研交流群
商业服务