遇见数据集

almanach/topxgen-deepseek-r1-distill-llama-70b-CoT

收藏
Hugging Face2025-10-05 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含了源文本、源语言、目标语言、翻译和目标文本等字段,被分为六个部分(CoT_T1至CoT_T6),每部分包含大约15万个示例。数据集主要用于文本翻译任务,具体语言对未在README中说明。

The dataset includes fields for source text, source language, target language, translation, and target text, and is divided into six parts (CoT_T1 to CoT_T6), each containing about 150,000 examples. The dataset is primarily for text translation tasks, but the specific language pairs are not specified in the README.

提供机构:
almanach
二维码
社区交流群
二维码
科研交流群
商业服务