遇见数据集

Replication Data for: Multi-label Prediction for Political Text-as-Data

收藏
Harvard Dataverse2021-04-07 更新2026-04-09 收录
官方服务:

资源简介:

Political scientists increasingly use supervised machine learning to code multiple relevant labels from a single set of texts. The current ``best practice'' of individually applying supervised machine learning to each label ignores information on inter-label association(s), and is likely to under-perform as a result. We introduce multi-label prediction as a solution to this problem. After reviewing the multi-label prediction framework, we apply it to code multiple features of (i) access to information requests made to the Mexican government and (ii) country-year human rights reports. We find that multi-label prediction outperforms standard supervised learning approaches, even in instances where the correlations among one's multiple labels are low. This repository replicates the figures and tables in the article and appendix. More information can be found in the "README.md" file.

创建时间:
2021-01-01
二维码
社区交流群
二维码
科研交流群
商业服务