遇见数据集

bismarck91/enA-frA-xc-one-tokenized-p3

收藏
Hugging Face2025-09-14 更新2025-10-25 收录
官方服务:

资源简介:

该数据集包含源代码和目标代码的整数序列对,适用于机器翻译或代码转换等任务。数据集共有200000个样本,分为训练集,文件大小为771096736字节,下载大小为289872684字节。

The dataset includes pairs of source and target code sequences represented as integer sequences, suitable for tasks such as machine translation or code conversion. The dataset consists of 200,000 samples in the training set, with a file size of 771096736 bytes and a download size of 289872684 bytes.

提供机构:
bismarck91
二维码
社区交流群
二维码
科研交流群
商业服务