This package contains data on five text analysis types (term extraction, contract analysis, topic modeling, network mapping), based on the survey data where researchers selected research output that a
Abstracts of reports describing clinical trials matched with study sample sizes extracted from ct.gov. To be used to train our sample size tagger: https://github.com/ijmarshall/robotlabs/blob/master/s
These are 30 sentences annotated by 15 crowd workers each, within the context of the project Crowd Watson (http://crowd-watson.nl) for medical relation extraction. Project members: Chris Welty (IBM Re