The data set is composed of 1494 txt files converted from CAJ format Chinese documents downloaded from CNKI, covering five topics: Bayesian network, personalized recommendation, image recognition, doc
Here we use misclassification errors as the evaluation metric. The best baseline performances are italicized, and the best overall performances are bolded.