llm87/semcor_tags
收藏官方服务:
资源简介:
该数据集包含多个字段,用于描述句子中的目标词及其相关信息。具体字段包括句子(sentence)、目标索引开始(target_index_start)、目标索引结束(target_index_end)、目标ID(target_id)、目标词元(target_lemma)、目标词性(target_pos)、意义键(sense_key)和标记句子(marked_sentence)。数据集仅包含一个训练集分割,共有218,806个示例,总大小为89,373,793字节。
The dataset includes multiple features such as sentence, target index start and end, target ID, target lemma, target POS, sense key, and marked sentence. The dataset is split into a training set containing 218806 samples. The download size of the dataset is 15754588 bytes, and the dataset size is 89373793 bytes.
提供机构:
llm87


