遇见数据集

Cross-Model Peer Review Dataset: AI Blind Reviews of The Force Multiplier Index

收藏
Zenodo2026-03-29 更新2026-05-26 收录
官方服务:

资源简介:

Complete audit trail of a cross-platform blind peer review process. Five frontier AI models (Claude Opus 4.6, ChatGPT 5.4, Gemini 3.1 Pro, Grok 4.2 Expert) independently reviewed The Force Multiplier Index using the same multi-dimensional evaluation prompt. The dataset contains 9 independent blind reviews, 5 cross-model syntheses, the review prompt, input and output documents, and NotebookLM-generated visual assets across two review rounds. Key findings: cross-model convergence on structural classification (70-80% agreement rate), a measurable spectrum of strictness in how models define parallel reasoning, unanimous identification of the same missing reasoning modes, and recursive self-correction when the framework was evaluated against its own quality criteria. See README.md for the complete file index, process map, and replication instructions.

提供机构:
Zenodo
创建时间:
2026-03-29
二维码
社区交流群
二维码
科研交流群
商业服务