遇见数据集

NeuralPGRank/hotpotqa-hard-negatives

收藏
Hugging Face2024-11-23 更新2024-12-14 收录
官方服务:

资源简介:

该数据集包含一组用于第二阶段重新排序的候选文档,这些候选文档由从gtr-t5-xl挖掘的硬负样本和已知与查询相关的真实文档组成。这是来自论文《Policy-Gradient Training of Language Models for Ranking》的发布,如果使用此数据集,请引用该论文。

This dataset contains a set of candidate documents for second-stage re-ranking on hotpotqa (test split in BEIR). Those candidate documents are composed of hard negatives mined from gtr-t5-xl as Stage 1 ranker and ground-truth documents that are known to be relevant to the query. This is a release from our paper Policy-Gradient Training of Language Models for Ranking, so please cite it if using this dataset.

提供机构:
NeuralPGRank
二维码
社区交流群
二维码
科研交流群
商业服务