遇见数据集

Supplementary Material: Impostor Phenomenon and Generative AI in Software Engineering -- A Systematic Literature Review

收藏
Zenodo2026-09-24 更新2026-10-01 收录
官方服务:

资源简介:

FILE DESCRIPTIONS 00_protocol_slr.pdf Full SLR protocol (version 2.0, 2026-05-13) in Spanish. Includes research questions, PICOC+F framework, inclusion/exclusion criteria (CI1-CI6, CE1-CE7), screening procedure, MMAT quality assessment plan, data extraction strategy (28 fields), and synthesis approach. Based on Kitchenham & Charters (2007) and PRISMA 2020 (Page et al., 2021). Note: CI2 was expanded on 2026-06-06 (director approval) to include adjacent constructs (self-efficacy, cognitive load, psychological ownership, intrinsic motivation, stress). 00_protocol_slr_en.pdf English translation of the full SLR protocol (version 2.0). 01_screening_decisions_F2.xlsx Phase 2 (full-text) screening decisions for all 57 studies evaluated at full text. Includes inclusion/exclusion verdict with textual evidence cited from each paper (section, page). Dual reviewer (R1 + R2) with arbitration by thesis director (R3) for disagreements. 27 excluded with documented reasons; 9 not retrieved with access attempts documented. 02_data_extraction_matrix.xlsx Data extraction matrix (v2) with 28 fields across 8 categories for all 24 included publications (23 studies). Categories: metadata, study design, population/demographics, IP measurement, GenAI interaction, productivity/team impact, interventions, limitations. E0216 and E0235 are extracted separately (different constructs reported) but counted as one study. 03_mmat_quality_scores.xlsx Mixed Methods Appraisal Tool (MMAT 2018) quality profiles for all 23 included studies. Each study scored on 2 screening items + 5 design-specific items as Yes/No/NPD (no global score, per MMAT 2018 guidelines). Dual evaluation by R1 and R2; inter-rater agreement: kappa = 0.49 (moderate, N=145 items). E0224 excluded post-extraction (MMAT annulled). 04_snowballing_tracking.xlsx Forward and backward snowballing tracking (Wohlin, 2014) applied to 14 seed studies across 3 tiers. Two directed expansion passes via Semantic Scholar (May-July 2026). 36 unique candidates evaluated; 4 incorporated (E0401-E0404). Second pass yielded one additional study (E0404). The traversal was not exhaustive: only two of 14 planned seeds and two secondary seeds were traversed. The 25 exclusions made solely by R0 were validated by R1 (25/25 confirmed, 0 discrepancies). The notas_r1 column contains the justification for each validation. Additional sheet "Validacion R0" with summary by exclusion criterion. 05_screening_decisions_F1.xlsx Phase 1 (title + abstract) consolidated screening decisions for all 269 unique records after deduplication. Includes AI agent (R0) advisory classification and independent human reviewer decisions (R1, R2). R1 and R2 independently evaluated 61 records; 54 agreements (88.5%); binary kappa (Include+Dudoso vs Exclude) = 0.80. R0 coincided with the final human classification in 109/262 records (41.6%; kappa = 0.114); R0 classified 234/269 as "Dudoso" (87.0%). 06_screening_agent.py Python script for the LLM-based screening agent (Llama 3.3 70B via Groq API). Advisory role only (R0); human reviewers made all final inclusion/exclusion decisions independently. Classifies records against CI/CE criteria and assigns relevance scores. The agent classified 87% of records as "Dudoso" (uncertain); its recall for inclusion was 100% (zero false negatives). 07_search_strings_final.txt Complete search strings (v4.2) for all 6 databases (Scopus, Web of Science, IEEE Xplore, EBSCO/PsycINFO, SpringerLink, ScienceDirect). Two complementary sub-chains: Foundation-A (IP x SE population) and GenAI-B (IP + adjacent constructs x SE x GenAI). Includes database-specific syntax adaptations. 08_kappa_calculation.py Python script for computing Cohen's kappa inter-rater agreement between human reviewers (R1 and R2). Supports both F1 screening and MMAT quality assessment agreement calculations. 09_exclusions_extraction_phase.md Decisions made after F2 screening during full-text extraction: (1) E0224 excluded for not measuring the impostor phenomenon; (2) E0239 included as adjacent construct (relative self-efficacy); (3) E0216/E0235 treated as one study with two publications. Documents the rationale and PRISMA implications for each decision. 10_screening_agent_procedure.md Detailed procedure for the AI-assisted screening (section 4.3): how each reviewer used the screening agent outputs, independence protocol, and the separation between R0 (advisory) and R1/R2 (decision) roles. 11_exclusiones_snowballing.md Resolution of the 6 snowballing candidates that remained as "Dudoso" after R0 screening. All 6 were reviewed by R1 on 2026-09-15 and excluded: 4 failed CI2 (construct not in the taxative list) and 2 failed CE4 (non-empirical/WIP). Includes rationale per case and final snowballing counts. 12_veredictos_consenso_mmat.pdf MMAT 2018 consensus verdicts document. Contains the agreed-upon quality profiles after three rounds of disagreement resolution between R1 and R2, verifying against original publications.

提供机构:
Zenodo
创建时间:
2026-09-24
二维码
社区交流群
二维码
科研交流群
商业服务