遇见数据集

BIOSSES (Biomedical Semantic Similarity Estimation System)

收藏
OpenDataLab2026-07-12 更新2024-05-09 收录
官方服务:

资源简介:

BIOSSES 数据集包含总共 100 个句子对,所有句子对均选自“TAC2 Biomedical Summarization Track Training Data Set”。句子对由五位不同的人类专家评估,他们判断它们的相似性并给出 [0-4] 范围内的分数。我们的指南是根据 SemEval 2012 任务 6 指南编写的。

The BIOSSES dataset contains a total of 100 sentence pairs, all selected from the "TAC2 Biomedical Summarization Track Training Data Set". These sentence pairs were evaluated by five distinct human experts, who judged their similarity and assigned scores within the range of [0, 4]. Our evaluation guidelines were developed based on the SemEval 2012 Task 6 guidelines.

提供机构:
OpenDataLab
创建时间:
2022-05-23
搜集汇总
数据集介绍
BIOSSES (Biomedical Semantic Similarity Estimation System) 数据集图片
背景与挑战
背景概述
BIOSSES是一个用于生物医学领域语义相似度估计的数据集,包含100个选自TAC2数据集的句子对,每个句子对由五位专家根据SemEval 2012指南标注0-4分的相似度分数。
以上内容由遇见数据集搜集并总结生成
二维码
社区交流群
二维码
科研交流群
商业服务