遇见数据集

The Cognitive Intern: A Design Study of Professional Judgment in Human-AI Collaboration

收藏
Zenodo2026-07-31 更新2026-08-01 收录
官方服务:

资源简介:

Qualitative Interviews, AI‑Generated Reports, and Credibility Annotations 1. Data Summary This dataset captures the first-use experiences of 42 domain-expert professionals in Pakistan with the platfom, a four-stage multi-agent system (MAS) pipeline. It provides a specialized resource for assessing how expert users judge AI credibility, localization, and technical depth in an emerging market context. 2. Data Composition The dataset is organized by Participant ID (P01–P42) and includes: 42 AI-Generated Reports (Reports/ folder): Business intelligence documents produced by the MAS pipeline. 42 Anonymized Transcripts (Transcripts/ folder): Clean text records of semi-structured interviews where experts evaluated the reports. Metadata (Metadata/metadata.csv): Details on the age, gender, professional domain, and years of experience for all 42 participants. 3. Annotation Definitions To ensure the analysis is reproducible, the following definitions were used to categorize expert feedback: Domain error: A factual or logical mistake identified by a participant using their specific professional expertise (e.g., P24 identifying the omission of "Chromite"). Localization failure: Missing, incorrect, or culturally insensitive information specific to the Pakistani context (e.g., incorrect tax rates or missing local landmarks). Note: While the broader study also evaluates Efficiency and Process Transparency, "Domain Error" and "Localization Failure" serve as the primary categorical labels for the machine learning benchmark tasks in this dataset. 4. Benchmarking Tasks This dataset supports the following research tasks: Credibility Prediction: Using expert transcripts to predict trust levels in specific AI-generated business outputs. Localization Quality Scoring: Measuring the accuracy of AI-generated cultural, legal, and economic content for non-Western regions. Process Transparency Analysis: Evaluating user disorientation during complex multi-agent reasoning phases. 5. Ethics & Privacy Informed Consent: All 42 participants provided explicit informed consent prior to the study. Anonymization: All transcripts and reports have been manually scrubbed of real names, company identities, and sensitive contact information. Institutional Oversight: This research was conducted at the Lab, XYZ University 6. Licensing This work is licensed under a Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0) license.

提供机构:
Zenodo
创建时间:
2026-07-31
二维码
社区交流群
二维码
科研交流群
商业服务