遇见数据集

bismarck91/enA-frA-xc-one-tokenized-p7

收藏
Hugging Face2025-09-15 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含源代码列表和目标代码列表两种整数序列类型的数据,主要用于训练代码生成或转换模型。数据集分为训练集,共有200000个示例,总文件大小为797032488字节。

The dataset includes source code lists and target code lists as integer sequence data, primarily used for training code generation or transformation models. The dataset is divided into a training set with a total of 200,000 examples and a total file size of 797,032,488 bytes.

提供机构:
bismarck91
二维码
社区交流群
二维码
科研交流群
商业服务