Accuracy results for four time spans.

Figshare2026-03-24 更新2026-04-28 收录

下载链接：

https://figshare.com/articles/dataset/_p_Accuracy_results_for_four_time_spans_p_/31846611

下载链接

链接失效反馈

官方服务：

资源简介：

Constructing datasets on past biodiversity from historical sources is crucial for understanding long-term ecological changes. Typically, compiling such datasets relies on prior knowledge of the sources’ composition and requires considerable manual effort. To overcome these challenges, we implement an automated approach based on prompted large language models (LLMs) to detect mentions of species in texts from 19th-century Württemberg and link these mentions to identifiers in the GBIF database. Based on our evaluation, we find that LLMs can reliably identify species in the texts with high recall (92.6%) and precision (95.3%), while providing estimates of the correct species identifier with considerable accuracy (83.0%). As our approach is easily scalable and adaptable to other contexts and languages, it offers a promising way to advance dataset generation from historical material using limited resources.

创建时间：

2026-03-24

5,000+

优质数据集

54 个

任务类型

进入经典数据集