stefanocarrera/autophagycode_D_he_train-mercury_Qwen3-8B_strategy_trust_t1.25_g1_run0_metrics
收藏资源简介:
该数据集包含代码任务执行结果和代码质量评估数据,用于软件工程和代码分析研究。特征包括任务ID、入口点、可执行性、正确性、测试通过和失败数量、错误类型,以及代码复杂度指标(如Halstead词汇量、长度、体积、难度、努力程度、时间)、圈复杂度、可维护性指数、代码行数(LOC和SLOC)、注释百分比、TTR(词型标记比)、令牌字典、香农熵、预测熵(均值和最大值)、定义函数数量和入口点重复标志。数据分割为训练集,包含164个示例,总大小约247KB。
This dataset encompasses code task execution outcomes and code quality evaluation data, tailored for software engineering and code analysis research. Its features include Task ID, Entry Point, Executability, Correctness, number of passed and failed tests, Error Types, code complexity metrics including Halstead Vocabulary, Length, Volume, Difficulty, Effort and Time, Cyclomatic Complexity, Maintainability Index, Lines of Code (LOC and SLOC), Comment Percentage, TTR (Type-Token Ratio), Token Dictionary, Shannon Entropy, Predictive Entropy (mean and maximum values), Number of Defined Functions, and Entry Point Duplication Flag. The dataset is split into a training set, which contains 164 examples with a total size of approximately 247 KB.




