登录后查看消息通知
搜索
常见问题
消息
登录
首页
/
数据集
/
Semantics_based_source_code_ labels.xlsx
Semantics_based_source_code_ labels.xlsx
收藏
NIAID Data Ecosystem
2026-03-11 收录
代码语义理解
源代码分类
数据链接:
https://figshare.com/articles/dataset/Semantics_based_source_code_labels_xlsx/12674291
数据链接
链接失效反馈
官方服务:
问题咨询
购买咨询
在线客服
NEW
资源简介:
This is the dataset of semantic labels from all labelled notebooks.
应用场景:
创建时间:
2020-07-20
相关数据集
Distribution shift datasets for source code classification
源代码分类
分布偏移
Please refer to https://github.com/testing-cs/CodeS for the description.
DataCite Commons
2025-04-01 更新
11
0
SCC++ classification result on `character count >= 26, line count >= 2`
源代码分类
代码分析
SCC classification result on `character count >= 26, line count >= 2`
DataCite Commons
2024-02-23 更新
10
0
Distribution shift datasets for source code classification
源代码分类
分布偏移
Three collections of datasets: Python75, Java250-S, Python800-S. Each collection has the same structure of directories. For example: Python75.zip: 1. raw: raw data files scrapped from the online resou
DataCite Commons
2022-08-25 更新
7
0
CodeT5SmallCAPS/CAPS_Java
代码语义理解
语言处理
--- configs: - config_name: default data_files: - split: train path: data/train-* dataset_info: features: - name: code dtype: string - name: code_sememe dtype: string - name: t
Hugging Face
2023-07-31 更新
10
0
AxiomicLabs/NPset-2-Python-Edu
代码语义理解
逻辑推理训练
NPset-2 (Python-Edu) 是一个标准化的半合成Python数据集,用于训练小语言模型在代码逻辑上,而无需处理原始代码语法的开销。它通过基于AST的转换器将Python源代码规范化,去除语法噪声,同时保留程序的完整逻辑结构,生成完全由自然语言标记组成的伪代码表示,使模型能够专注于代码的功能而非形式。
Hugging Face
2026-05-07 更新
8
0
© 2023-2026 上海数据发展科技有限责任公司 版权所有
沪ICP备17003045号-15
沪公网安备31010402336585号
热门搜索
社区交流群
科研交流群
商业服务
数据资源
寻源服务
数据采集
标注服务
数据产品
代理销售
数据领域
凭证登记
数据产品
介绍推广