A Systematic Evaluation of Large Language Models of Code
收藏数据链接:
官方服务:
资源简介:
Test sets of ~100 files in each of 12 programming languages. These files are not included in The Pile, and thus models such as GPT-Neo, GPT-J, GPT-NeoX were not trained on them. In the paper, we use these test sets to compare a variety of language models of code including OpenAI's Codex, GPT-J, GPT-Neo, GPT-NeoX-20B, and CodeParrot and our PolyCoder model.
提供机构:
Zenodo创建时间:
2022-03-08



