The Sarah AI Evaluation Corpus v1.0 (Dataset for "CAG-4S")
收藏资源简介:
This dataset contains the supplementary materials for the paper "CAG-4S: A 4-Pillar, Pure-Context Memory Architecture for Emergent Narrative Identity in LLM Agents." https://doi.org/10.5281/zenodo.17391038 The corpus serves as the primary empirical evidence for the claims made in the paper and includes the following files: interview_intro.jsonl: The complete, unedited, machine-readable transcript of the live evaluation interview demonstrated in the accompanying video. interview_main_18Q.jsonl: The 18-question dialogue. evaluation_questions.md: A human-readable transcript of the final 18 meta-cognitive questions. sarah.mp4: The final, uncut video demonstration. README.md: Contains metadata, licensing information (CC-BY 4.0), and the SHA-256 checksum for the video file. This dataset allows for the full replication and analysis of the evaluation presented in the paper.
本数据集为论文《CAG-4S:面向大语言模型智能体(LLM Agents)的涌现叙事身份的四支柱纯上下文记忆架构》的补充材料。 对应的DOI链接:https://doi.org/10.5281/zenodo.17391038 该语料库为本论文提出的各项主张提供核心实证依据,包含以下文件: interview_intro.jsonl:配套视频中展示的现场评估访谈的完整未编辑机器可读笔录 interview_main_18Q.jsonl:含18个问题的对话语料 evaluation_questions.md:最终18个元认知问题的人类可读笔录 sarah.mp4:完整未剪辑的最终演示视频 README.md:包含元数据、许可协议信息(CC-BY 4.0)以及该视频文件的SHA-256校验和 本数据集可实现论文中所述评估实验的完整复现与分析。



