Wordsim-240
收藏官方服务:
资源简介:
wordsim-240数据集是一个词向量数据集,向量表示每个词的句法和语义信息,可以用来解决各种NLP任务。该数据集提供了手动注释的中文词汇对和相似度分数,它们从相应的英文数据翻译成中文数据。
The Wordsim-240 dataset is a word vector dataset that encodes the syntactic and semantic information of individual words, and can be used to address various natural language processing (NLP) tasks. This dataset provides manually annotated Chinese word pairs and their similarity scores, which are translated from the corresponding English dataset.
提供机构:
OpenDataLab创建时间:
2023-04-20
搜集汇总
数据集介绍

背景与挑战
背景概述
Wordsim-240是一个词向量数据集,通过向量表示词汇的句法和语义信息,可用于多种自然语言处理任务。该数据集提供了从英文翻译而来的中文词汇对及其手动标注的相似度分数。
以上内容由遇见数据集搜集并总结生成



