This dataset includes responses from 1,049 participants from 16 districts in Beijing, between September 8 and 19, 2022: 352 in the Group-by-Case, 351 in the Group-by-District, and 346 in the Group-by-
Validation data for the Astro scientific publication clustering benchmark dataset This is the dataset used in the publication Donner, P. "Validation of the Astro dataset clustering solutions with ex
this is the raw data of the manuscript "a new hybrid citation-text model for scientific document clustering." The data contain two parts. One is the files in xml format, which are download from PMC da
The PubMed model contains over 18 million PubMed documents (1996-2019) clustered into 28,743 clusters for use in research planning, portfolio analysis, systematic review, etc. This repository contains