遇见数据集

QualDeltaBench: A Review-Grounded Benchmark for Evaluating LLMs as Code Quality Judges

收藏
Zenodo2026-07-08 更新2026-08-02 收录
官方服务:

资源简介:

This artifact contains the compressed experimental outputs produced by the evaluated models, with one .zip archive per model. Each archive includes the results for the supported experimental settings, including function-scoped, file-scoped, and repository-scoped runs, with both aspect-conditioned and open-aspect configurations on both Java and Python datasets.

提供机构:
Zenodo
创建时间:
2026-07-08
二维码
社区交流群
二维码
科研交流群
商业服务