遇见数据集

deshanksuman/WSD_DATASET_FEWS_SEMCOR

收藏
Hugging Face2025-03-21 更新2025-04-12 收录
官方服务:

资源简介:

FEWS和Semcor词义消歧数据集是一个经过预处理和格式化的数据集,旨在直接用于训练和微调用于词义消歧的语言模型。数据集中的每个模糊词都被特殊标签`<WSD>`包围,以便模型在训练和推理时专注于这些特定的词。

The FEWS and Semcor Dataset for Word Sense Disambiguation (WSD) is a preprocessed and formatted dataset designed to be directly used for training and fine-tuning language models for word sense disambiguation. Each ambiguous word in the dataset is enclosed with `<WSD>` tags to focus the model on these specific words during training and inference.

提供机构:
deshanksuman
二维码
社区交流群
二维码
科研交流群
商业服务