Topic Extraction Dataset
收藏官方服务:
资源简介:
In this dataset, a total of 9691 articles within the medical domain were collected for analysis. Topic extraction was conducted utilizing two distinct methodologies: Textrank and LLM. These approaches were applied in conjunction with the keywords present in the articles, forming the dataset for analysis. The dataset encompasses various fields such as article title, publication year, PMID, keyword listings, topics derived through the Textrank algorithm, and topics identified through LLM.
创建时间:
2024-04-03



