Here are the datasets used for GROBID end-to-end benchmarking covering: - metadata extraction, - bibliographical reference extraction, parsing and citation context identification, and - full tex
Additional file 4 Supplementary Table 2. GWAS phenotypes parsed by Nelson’s group and pyMeSHSim, and the semantic similarity between them calculated by pyMeSHSim and meshes.