遇见数据集

Rediscovery Datasets: Connecting Duplicate Reports Of Apache, Eclipse, And Kde

收藏
Zenodo2020-09-18 更新2026-05-25 收录
数据链接:
官方服务:

资源简介:

We present three defect rediscovery datasets mined from Bugzilla. The datasets capture data for three groups of open source software projects: Apache, Eclipse, and KDE. The datasets contain information about approximately 914 thousands of defect reports over a period of 18 years (1999-2017) to capture the inter-relationships among duplicate defects. <strong>File Descriptions</strong> apache.csv - Apache Defect Rediscovery dataset eclipse.csv - Eclipse Defect Rediscovery dataset kde.csv - KDE Defect Rediscovery dataset apache.relations.csv - Inter-relations of rediscovered defects of Apache eclipse.relations.csv - Inter-relations of rediscovered defects of Eclipse kde.relations.csv - Inter-relations of rediscovered defects of KDE create_and_populate_neo4j_objects.cypher - Populates Neo4j graphDB by importing all the data from the CSV files. Note that you have to set dbms.import.csv.legacy_quote_escaping configuration setting to false to load the CSV files as per https://neo4j.com/docs/operations-manual/current/reference/configuration-settings/#config_dbms.import.csv.legacy_quote_escaping create_and_populate_mysql_objects.sql - Populates MySQL RDBMS by importing all the data from the CSV files rediscovery_db_mysql.zip - For your convenience, we also provide full backup of the MySQL database neo4j_examples.txt - Sample Neo4j queries mysql_examples.txt - Sample MySQL queries rediscovery_eclipse_6325.png - Output of Neo4j example #1 distinct_attrs.csv - Distinct values of bug_status, resolution, priority, severity for each project

提供机构:
Zenodo
创建时间:
2017-03-18
二维码
社区交流群
二维码
科研交流群
商业服务