遇见数据集

chizk/Wikipedia-zh-tw

收藏
Hugging Face2026-05-08 更新2026-05-31 收录
官方服务:

资源简介:

该数据集包含2,533,212个训练示例,每个示例由question(问题)和answer(答案)两个文本字段组成,数据格式为字符串。数据集总大小约为8.56GB,下载大小约为3.68GB,适用于自然语言处理任务,如问答系统训练或对话生成,但具体内容、来源或应用场景未在README中说明。

This dataset contains 2,533,212 training examples, each consisting of two text fields: question and answer, both in string format. The total dataset size is approximately 8.56GB, with a download size of about 3.68GB. It is suitable for natural language processing tasks such as question-answering system training or dialogue generation, but specific content, sources, or application scenarios are not described in the README.

提供机构:
chizk
二维码
社区交流群
二维码
科研交流群
商业服务