遇见数据集

LAGT

收藏
Zenodo2024-10-04 更新2026-06-04 收录
数据链接:
官方服务:

资源简介:

LAGT is s a dataset of lemmatized ancient Greek texts, combining works from Perseus Digital Library and First 1000 Years of Greek. The scripts used to produce this version of the dataset are available from //github.com/sdam-au/LAGT. Concerning lemmatization, the dataset contains lemmatized sentences in a form of list-of-lists, with sublist elements representing individual lemmata. It contains only nouns, proper names, verbs and adjectives. Whenever available, the lemmata are based on the GLAUx corpus: https://github.com/perseids-publications/glaux-trees.

提供机构:
Zenodo
创建时间:
2022-10-18
二维码
社区交流群
二维码
科研交流群
商业服务