deshanksuman/WSD_DATASET_FEWS_SEMCOR
收藏官方服务:
资源简介:
FEWS和Semcor词义消歧数据集是一个经过预处理和格式化的数据集,旨在直接用于训练和微调用于词义消歧的语言模型。数据集中的每个模糊词都被特殊标签`<WSD>`包围,以便模型在训练和推理时专注于这些特定的词。
The FEWS and Semcor Dataset for Word Sense Disambiguation (WSD) is a preprocessed and formatted dataset designed to be directly used for training and fine-tuning language models for word sense disambiguation. Each ambiguous word in the dataset is enclosed with `<WSD>` tags to focus the model on these specific words during training and inference.
提供机构:
deshanksuman


