数据链接:
官方服务:
资源简介:
Dataset for comparison of trends of code analysis tools from 2004 to 2023.
应用场景:
创建时间:
2023-04-16
相关数据集
autophagycode_D_metrics_he_Qwen3-8B_lr0.0001_correct_g3
该数据集包含与代码执行和质量评估相关的结构化数据,主要用于代码分析、缺陷检测和软件质量评估任务。数据集包含164个训练样本,每个样本包含14个特征字段:任务ID(task_id)、入口函数(entry_point)、可执行状态(is_executable)、正确性标识(is_correct)、通过测试数(tests_passed)、失败测试数(tests_failed)、测试运行时间(test_r
Hugging Face2026-03-24 更新150
lycfight/lycfight__SWE-bench_Verified_0__style-2__fs-oracle
这个数据集包含了软件仓库中的代码问题和补丁信息,每个记录包括实例ID、文本描述、代码仓库信息、基础提交哈希值、问题描述、提示文本、创建时间、补丁代码、测试补丁代码、版本号以及两个状态字段(FAIL_TO_PASS和PASS_TO_PASS),表示代码从失败到通过或从通过到通过的状态变化。数据集仅包含一个测试集,大小为23288字节。
Hugging Face2025-04-09 更新110
Comparing Maintainability Index, SIG Method, and SQALE for Technical Debt Identification
Comparing Maintainability Index, SIG Method, and SQALE for Technical Debt Identification This repository contains the mined datasets, raw data analysis, graphical outputs, and source code used for the
NIAID Data Ecosystem60
reshinthadith/2048_has_code_filtered_base_code_review_python
--- dataset_info: features: - name: body dtype: string - name: comments list: - name: ContentLicense dtype: string - name: CreationDate dtype: string - name: Id
Hugging Face2023-01-19 更新150
xszheng2020/the_stack_dedup_python_hits_1_qsc_code_cate_autogen
这是一个关于代码质量评估的数据集,包含了多种编程语言编写的代码文件的特性,如文件哈希值、大小、扩展名、代码内容统计特征等。数据集还包含了仓库的星星数、问题数和分支数等信息,以及相关的时间戳。这些数据可能用于训练评估代码质量的机器学习模型。
Hugging Face2025-09-20 更新100



