QualDeltaBench: A Review-Grounded Benchmark for Evaluating LLMs as Code Quality Judges
收藏官方服务:
资源简介:
This artifact contains the compressed experimental outputs produced by the evaluated models, with one .zip archive per model. Each archive includes the results for the supported experimental settings, including function-scoped, file-scoped, and repository-scoped runs, with both aspect-conditioned and open-aspect configurations on both Java and Python datasets.
提供机构:
Zenodo创建时间:
2026-07-08



