Wordsim-297
收藏官方服务:
资源简介:
wordsim-297数据集是一个词向量数据集,向量表示每个词的句法和语义信息,可以用来解决各种NLP任务。该数据集提供了手动注释的中文词汇对和相似度分数,它们从相应的英文数据翻译成中文数据。
The WordSim-297 dataset is a word vector dataset that encodes the syntactic and semantic information of individual words through dense vector representations, and can be leveraged to address a wide range of natural language processing (NLP) tasks. It includes manually annotated Chinese word pairs and their corresponding similarity scores, which are translated from the matching English-language datasets.
提供机构:
OpenDataLab创建时间:
2023-04-20
搜集汇总
数据集介绍

背景与挑战
背景概述
Wordsim-297是一个词向量数据集,通过向量表示词的句法和语义信息,适用于多种自然语言处理任务。该数据集提供了从英文翻译而来的中文词汇对及其手动标注的相似度分数。
以上内容由遇见数据集搜集并总结生成



