ManyTypes4Py: A Benchmark Python Dataset for Machine Learning-Based Type Inference
收藏NIAID Data Ecosystem2026-03-12 收录
下载链接:
https://zenodo.org/record/4044635
下载链接
链接失效反馈官方服务:
资源简介:
The dataset is gathered on Sep. 17th 2020 from GitHub.
It has clean and complete versions (from v0.7):
The clean version has 5.1K type-checked Python repositories and 1.2M type annotations.
The complete version has 5.2K Python repositories and 3.3M type annotations.
The dataset's source files are type-checked using mypy (clean version).
The dataset is also de-duplicated using the CD4Py tool.
Check out the README.MD file for the description of the dataset.
Notable changes to each version of the dataset are documented in CHANGELOG.md.
The dataset's scripts and utilities are available on its GitHub repository.
创建时间:
2021-08-25



