nltk-data-hub/stopwords
收藏官方服务:
资源简介:
NLTK停用词数据集包含来自NLTK的停用词列表,涵盖33种语言。每种语言都是一个独立的配置,每行代表一个停用词。该数据集适用于文本分类和标记分类任务。
The NLTK Stopwords dataset contains stopword lists from NLTK, covering 33 languages. Each language is a separate config, and each row represents one stopword. The dataset is intended for text-classification and token-classification tasks.
提供机构:
nltk-data-hub


