The data file contains a list of articles and their RCT Tagger prediction scores, which were used in a project associated with the manuscript "Evaluation of an automated probabilistic RCT Tagger appli
These datasets are derived from the data provided and originally used by: Cohen A.M., Hersh W.R., Peterson K., Yen P.-Y. (2006): Reducing workload in systematic review preparation using automated cita
In order to improve the move recognition performance, We construct a refined corpus based on PubMed called RCMR 280k. The corpus consists of approximately 280,000 structured abstracts, totaling 3,386,