Prompt and Configuration Package for Generative-AI Harness Evaluation in Multilingual School Communication Support
收藏资源简介:
本資料は,多言語学校コミュニケーション支援に生成AIを導入する際の支援形態別プロンプトおよび公開用設定をまとめたものです。 本資料には,5つの支援形態に対応する候補出力生成プロンプト,補助judgeプロンプト,共通system prompt,出力スキーマ注記,実験設定・モデル群・judge設定・採点方針に関する公開用configが含まれます。 本資料は実行コード一式ではありません。評価コード,API実行コード,provider adapter,評価ケース,モデル出力,人手レビュー記録,集計済みアウトプット,ヒアリング生ログ,個人情報は含みません。論文で用いた支援形態設計と公開用評価設定の透明性・監査可能性を高めることを目的としています。 補助judgeプロンプトは,人手レビューを置き換えるものではなく,人手レビューが必要な出力を抽出する補助として位置づけています。 ===== This record provides prompt templates and public configuration files for evaluating generative-AI support modes in multilingual school communication support, with attention to language young carer burdens. The package includes candidate-generation prompts for five support modes, auxiliary LLM-as-a-judge prompts, a common system prompt, output schema notes, and public configuration files for experiment settings, model groups, judge settings, and scoring policy. This is not an executable code package and does not include evaluation code, API execution code, provider adapters, evaluation cases, model outputs, human-review records, aggregate result files, raw interview logs, or personally identifiable information. The purpose of this release is to support transparency and auditability of the prompt design and public evaluation settings used in the related paper. The LLM-as-a-judge prompts are intended as auxiliary tools for identifying outputs that require human review. They are not intended to replace human evaluation.



