AI paper index
Pragmatic Task Battery: LLM Responses and Rubric Scores
One-line summary
An AI research paper on Pragmatic Task Battery: LLM Responses and Rubric Scores.
Engineering notes
Engineering notes will be added by the aipentium editorial team.
Chinese explanation / 中文解读
中文解读待补充:本站会优先为大语言模型、生成式AI、ChatGPT相关技术、计算机视觉、深度学习等高价值论文补充中文说明。
Original abstract
This dataset accompanies the manuscript "Pragmalinguistic competence without sociopragmatic calibration: an exploratory assessment of pragmatic performance in four large language models," submitted to the Journal of Pragmatics. It contains the complete Pragmatic Task Battery (PTB) — twelve Discourse Completion Task (DCT) scenarios — the verbatim responses produced by four large language models (ChatGPT / GPT-5.3 Instant, Gemini 3 Flash, Claude Sonnet 4.6, and Perplexity Sonar Pro) to each scenario, and the Python script used to compute the statistical analysis reported in the manuscript (Friedman test, Kendall's W, post-hoc Wilcoxon signed-rank comparisons).
Links and sources
Need this topic turned into a technical roadmap?
aipentium can prepare a custom AI literature review, code map, dataset map, and B2B technology assessment.
Request B2B AI research
Comments