遇见数据集

llm87/semcor_tags

收藏
Hugging Face2024-12-05 更新2024-12-14 收录
官方服务:

资源简介:

该数据集包含多个字段,用于描述句子中的目标词及其相关信息。具体字段包括句子(sentence)、目标索引开始(target_index_start)、目标索引结束(target_index_end)、目标ID(target_id)、目标词元(target_lemma)、目标词性(target_pos)、意义键(sense_key)和标记句子(marked_sentence)。数据集仅包含一个训练集分割,共有218,806个示例,总大小为89,373,793字节。

The dataset includes multiple features such as sentence, target index start and end, target ID, target lemma, target POS, sense key, and marked sentence. The dataset is split into a training set containing 218806 samples. The download size of the dataset is 15754588 bytes, and the dataset size is 89373793 bytes.

提供机构:
llm87
二维码
社区交流群
二维码
科研交流群
商业服务