遇见数据集

Wiki-CS Dataset

收藏
paperswithcode.com2025-01-21 收录
官方服务:

资源简介:

Wiki-CS is a Wikipedia-based dataset for benchmarking Graph Neural Networks. The dataset is constructed from Wikipedia categories, specifically 10 classes corresponding to branches of computer science, with very high connectivity. The node features are derived from the text of the corresponding articles. They were calculated as the average of pretrained GloVe word embeddings (Pennington et al., 2014), resulting in 300-dimensional node features. The dataset has 11,701 nodes and 216,123 edges.

Wiki-CS 是一项基于维基百科的用于评估图神经网络性能的数据集。该数据集由维基百科的类别构建而成,具体包括对应计算机科学各分支的10个类别,具有极高的连通性。节点特征源自相应文章的文本内容,经过计算得出,作为预训练的 GloVe 词嵌入(Pennington 等人,2014年)的平均值,从而形成300维的节点特征。该数据集包含11,701个节点和216,123条边。

提供机构:
Papers with Code
搜集汇总
背景与挑战
背景概述
Wiki-CS是一个基于Wikipedia构建的图神经网络基准测试数据集,包含10个计算机科学分支类别,具有高连通性。节点特征来自文章文本,通过预训练GloVe词嵌入平均计算为300维,数据集规模为11,701个节点和216,123条边。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务