遇见数据集

sara0123456789/GenAI-Dataset

收藏
Hugging Face2026-05-10 更新2026-05-31 收录
官方服务:

资源简介:

该数据集包含三个划分(训练集、验证集和测试集),总大小约为74.8 MB,共有207,496个示例。每个示例由两个字符串类型的特征组成:input(输入)和output(输出)。训练集有165,996个示例,验证集和测试集各有20,750个示例。数据集可能用于自然语言处理任务,如文本生成或转换,但具体领域和用途未在README中明确说明。

This dataset consists of three splits (train, validation, and test), with a total size of approximately 74.8 MB and 207,496 examples. Each example includes two string features: input and output. The train split has 165,996 examples, while both validation and test splits have 20,750 examples each. The dataset may be intended for natural language processing tasks, such as text generation or transformation, but specific domain and usage are not explicitly mentioned in the README.

提供机构:
sara0123456789
二维码
社区交流群
二维码
科研交流群
商业服务