vaibhavalakshmiravideshik/mesh-snomed-entity-alignment-15k
收藏资源简介:
MeSH-SNOMED Entity Alignment 15K 是一个生物医学异质知识图谱对齐基准数据集,用于在[MeSH](https://www.nlm.nih.gov/mesh/)和[SNOMED CT](https://www.snomed.org/what-is-snomed-ct)之间进行跨本体匹配。该数据集旨在评估实体对齐系统在现实大图条件下的性能,其中黄金对齐的概念嵌入在包含许多结构相关但非对齐背景实体的更大生物医学图谱中。此版本为伴随的**EMNLP 2026**提交而设计。基准数据集从现有的MeSH-SNOMED匹配集合开始,并将这些对齐与从两个源本体提取的图谱上下文打包在一起。数据集的目的是不是重新判定每个原始匹配的语义有效性,而是支持研究对齐系统是否能够在大型异质图谱空间中恢复语义对应的生物医学实体。
MeSH-SNOMED Entity Alignment 15K is a biomedical heterogeneous knowledge graph alignment benchmark for cross-ontology matching between [MeSH](https://www.nlm.nih.gov/mesh/) and [SNOMED CT](https://www.snomed.org/what-is-snomed-ct). It is designed to evaluate entity alignment systems under realistic large-graph conditions, where gold-aligned concepts are embedded in much larger biomedical graphs containing many structurally relevant but non-aligned background entities. This release is intended for the accompanying **EMNLP 2026** submission. The benchmark starts from an existing set of MeSH-SNOMED matches and packages those alignments together with graph context extracted from the two source ontologies. The purpose of the dataset is not to newly adjudicate the semantic validity of every original match, but to support research on whether alignment systems can recover semantically corresponding biomedical entities across large, heterogeneous graph spaces.




