相关数据集
BSpell: A CNN-Blended BERT Based Bengali Spell Checker Dataset
Bengali typing is mostly performed using English keyboard and can be highly erroneous due to the presence of compound and similarly pronounced letters. Spelling correction of a misspelled word require
Zenodo2023-02-24 更新20
Bug-fix commit script and dataset
This contains the scripts and dataset used in "automatic identification of bug-fixing commits in Python open-source projects" for reproduction purposes.
Zenodo2021-01-15 更新10
Nutrients measured on water bottle samples at station DI191_12095#3
Plymouth Marine Laboratory. Authorship was originally taken from the JGOFS CDs (Lowry & BODC) and was changed by request of the BODC.
PANGAEA2004-01-01 更新00
基于思维链的软件漏洞自动修复技术研究
环境需求: 请安装以下程序以满足CotRepair的基本运行需求: pytorch=2.0.0; datasets==1.16.1; transformers==4.21.1; nltk=3.8.1; scipy=1.10.1; 结果复现: 使用如下命令可以训练CotRepair:
Zenodo2023-07-04 更新10
ANNalog — Generation of MedChem-similar Molecules
source SMILES are saved in .src files, and target SMILES are saved in .trg files when training the model, use train.src and train.trg for seq2seq model training, the first SMILES from .src file and f
Zenodo2025-09-18 更新00



