stefanocarrera/autophagycode_D_he_train-mercury_Qwen3-4B_strategy_trust_t1.25_g3_run0_metrics
收藏资源简介:
该数据集包含编程任务相关的元数据,用于代码分析和质量评估。特征包括任务ID、入口点、可执行性、正确性、测试通过/失败数量、错误类型,以及多种代码复杂度指标(如Halstead词汇量、长度、体积、难度、努力程度和时间,圈复杂度,可维护性指数),代码统计信息(如代码行数、有效代码行数、注释比例),语言多样性指标(TTR、Shannon熵),预测熵值,以及函数定义数量等。数据集可能用于机器学习模型训练,以评估代码质量、自动化测试或代码生成任务。
This dataset contains metadata related to programming tasks, designed for code analysis and quality assessment. Features include task ID, entry point, executability, correctness, number of tests passed/failed, error type, and various code complexity metrics (e.g., Halstead vocabulary, length, volume, difficulty, effort, and time, cyclomatic complexity, maintainability index), code statistics (e.g., lines of code, source lines of code, comment percentage), language diversity metrics (TTR, Shannon entropy), predictive entropy values, and number of function definitions. The dataset is likely used for training machine learning models to evaluate code quality, automate testing, or code generation tasks.




