GPT4Kids Research Data: LLM-Generated Linguistic Corpora for German Children's Word Frequency Estimation
收藏官方服务:
资源简介:
Dataset for "Can Large Language Models generate useful linguistic corpora? A case study of the word frequency effect in young German readers" Contains data from three experiments: (1) GPT-3.5 (2) temperature and audience variation analysis, and (3) open-weight LLMs evaluation. Keywords: psycholinguistics, word frequency, language models, children reading, German corpus linguistics, experimental data, GPT-3.5, Llama, DeepSeek See https://github.com/jobschepens/gpt4kids
提供机构:
Zenodo创建时间:
2025-07-24



