遇见数据集

Automatic financial term extractor

收藏
DataCite Commons2025-11-12 更新2025-04-10 收录
官方服务:

资源简介:

<p>The creation of this dataset is framed in the Spanish national project CLARA-FINT. The aim of this task within the project was to create an automatic financial term extractor for Spanish. In order to do so, the first step was to apply linguistic annotation on texts, namely annual reports from the main Spanish listed companies in the IBEX 35 index. The next step involved the use of these annotations to fine-tune a model for the financial term extraction task. This dataset contains the fine-tuned model, i.e., the automatic extractor. It is described in the paper PORTA-ZAMORANO, J., CARBAJO-CORONADO, B., MORENO-SANDOVAL, A. (2024).</p> <p>It is a bert-multilingual model that was fine-tuned for the financial terms extraction task. Texts annotated by linguists were used for fine-tuning as training data. Said texts contained financial terms that were highlighted within their context.</p>

提供机构:
e-cienciaDatos
创建时间:
2025-03-13
二维码
社区交流群
二维码
科研交流群
商业服务