stefanocarrera/autophagycode_D_he_train-mercury_Qwen3-4B_strategy_trust_t1.25_g10_run0_metrics
收藏资源简介:
该数据集包含代码任务相关的执行和测试数据,用于分析和评估代码质量。特征包括任务标识、入口点、可执行状态、正确性、通过和失败的测试数量、测试运行时间、错误类型,以及Halstead度量(如词汇量、长度、体积、难度、努力程度、时间)、圈复杂度、可维护性指数、代码行数(LOC和SLOC)、注释百分比、TTR(类型标记比)、令牌字典、香农熵、预测熵(均值和最大值)、定义函数数量、入口点重复性等。数据集适用于代码复杂度分析、可维护性评估、自动化测试和机器学习模型训练。
This dataset contains execution and testing data related to code tasks, designed for analyzing and evaluating code quality. Features include task ID, entry point, executable status, correctness, number of tests passed and failed, test run time, error type, as well as Halstead metrics (e.g., vocabulary, length, volume, difficulty, effort, time), cyclomatic complexity, maintainability index, lines of code (LOC and SLOC), comment percentage, TTR (type-token ratio), token dictionary, Shannon entropy, predictive entropy (mean and maximum), number of functions defined, and entry point repetition. The dataset is suitable for code complexity analysis, maintainability assessment, automated testing, and machine learning model training.




