遇见数据集

SequenceR data new

收藏
Zenodo2021-05-06 更新2026-05-28 收录
数据链接:
官方服务:

资源简介:

Dataset of single line patches for SequenceR. Obtained by preprocessing and merging datasets CodRep and the one used for Learning Bug-Fixing Patches in the Wild via Neural Machine Translation. Preprocessing steps will be located at https://github.com/KTH/chai/tree/master/src/embedding-work The structure of the tar file is as follows: sequencer-data.tar.gz<br> ├── test<br> │ ├── meta.tsv<br> │ ├── 1<br> │ │ ├── buggy_file.java<br> │ │ └── fixed_line.txt<br> │ ├── 2<br> │ │ ├── buggy_file.java<br> │ │ └── fixed_line.txt<br> │ ├── 3<br> │ │ ├── ...<br> ├── train<br> │ ├── meta.tsv<br> │ ├── 1<br> │ │ ├── ...<br> │ │ ...<br> .java files contain the abstracted source code (see SequenceR paper), .txt files contain the fixed line, and the meta.tsv files contain a mapping between each dir (1, 2, ...) and a unique id of the example, and the line number for the buggy line.

提供机构:
Zenodo
创建时间:
2021-05-06
二维码
社区交流群
二维码
科研交流群
商业服务