官方服务:
资源简介:
Webservice for Weblicht
面向Weblicht的Web服务
应用场景:
相关数据集
SEACrowd/khmer_alt_pos
该数据集包含20,000个高棉语句子,这些句子经过了手动分词和词性标注(POS-tagging)的注释。数据集的主要用途是支持词性标注任务。数据集的语言为高棉语(khm)。
Hugging Face2024-06-24 更新230
prakod/GLuecos_POS_EN_HI_FG
--- dataset_info: features: - name: words sequence: string - name: label1 sequence: string - name: label2 sequence: string splits: - name: dev_Romanized num_bytes: 110280
Hugging Face2024-05-31 更新160
Lots-of-LoRAs/task1167_penn_treebank_coarse_pos_tagging
该数据集名为task1167_penn_treebank_coarse_pos_tagging,属于文本生成任务类别,主要用于自然语言处理中的粗粒度词性标注任务。数据集包含训练集、验证集和测试集,分别有5184、648和648个样本。每个样本包含输入、输出和ID三个特征。数据集的创建者和语言创建者均为众包,语言为英语,许可证为Apache-2.0。数据集的相关研究可以在提供的论文中找到,详细信息请
Hugging Face2024-07-16 更新80
Buckwalter Arabic Morphological Analyzer Version 2.0
Introduction This file contains documentation on the Buckwalter Arabic Morphological Analyzer Version 2.0. Data The data consists primarily of three Arabic-E
DataCite Commons2021-07-01 更新220
Noun Verb Dataset
该数据集包含自然发生的英语句子,这些句子具有非平凡的名词-动词歧义。数据集用于帮助英语词性标注器改进在名词-动词歧义上的表现,从而提高翻译和文本到语音合成的下游任务的准确性。
github2024-02-26 更新140



