Here we use misclassification errors as the evaluation metric. The best baseline performances are italicized, and the best overall performances are bolded.
The data set is composed of 1494 txt files converted from CAJ format Chinese documents downloaded from CNKI, covering five topics: Bayesian network, personalized recommendation, image recognition, doc