Pragmatic Task Battery: LLM Responses and Rubric Scores
收藏Mendeley Data2026-09-09 收录
官方服务:
资源简介:
This dataset accompanies the manuscript "Pragmalinguistic competence without sociopragmatic calibration: an exploratory assessment of pragmatic performance in four large language models," submitted to the Journal of Pragmatics. It contains the complete Pragmatic Task Battery (PTB) — twelve Discourse Completion Task (DCT) scenarios — the verbatim responses produced by four large language models (ChatGPT / GPT-5.3 Instant, Gemini 3 Flash, Claude Sonnet 4.6, and Perplexity Sonar Pro) to each scenario, and the Python script used to compute the statistical analysis reported in the manuscript (Friedman test, Kendall's W, post-hoc Wilcoxon signed-rank comparisons).
创建时间:
2026-08-18




