遇见数据集

Supplementary Material: An Exploratory Study on Prompting Strategies for LLM-Assisted Heuristic Evaluation of Web User Interfaces

收藏
Zenodo2026-08-11 更新2026-08-13 收录
官方服务:

资源简介:

User Interfaces (UIs) — The 12 AI-generated e-commerce web interfaces used as evaluation targets in the study. Dataset — The set of Nielsen heuristic violations identified for each interface, comprising both the LLM-generated evaluations (GPT under zero-shot, few-shot, and chain-of-thought prompting) and the ground-truth annotations produced by a senior usability specialist, used as the basis for comparison. Analysis Code — A Jupyter notebook (.ipynb) implementing the full analysis pipeline, including the computation of agreement metrics (F1-score and Cohen's Kappa) between the LLM-generated evaluations and the specialist's ground truth across the three prompting strategies.

提供机构:
Zenodo
创建时间:
2026-08-11
二维码
社区交流群
二维码
科研交流群
商业服务