遇见数据集

timchen0618/browsecomp-plus-benchmark

收藏
Hugging Face2026-05-26 更新2026-05-31 收录
官方服务:

资源简介:

该数据集是一个问答或信息检索相关的数据集,包含830个训练样本。每个样本包括查询ID(query_id)、查询内容(query)、答案(answer)、证据文档(evidence_docs)和黄金文档(gold_docs)等字段,所有字段均为字符串类型。证据文档和黄金文档可能用于支持答案的验证或评估。数据集仅提供训练集,总大小为157129761字节,下载大小为83582343字节。基于字段命名,推测该数据集可能用于训练或评估问答系统、文档检索或自然语言处理任务,但具体应用和背景未在README中说明。

This dataset is related to question answering or information retrieval, containing 830 training examples. Each example includes fields such as query_id, query, answer, evidence_docs, and gold_docs, all of which are of string type. Evidence documents and gold documents may be used for answer validation or evaluation. The dataset only provides a training split, with a total size of 157129761 bytes and a download size of 83582343 bytes. Based on the field names, it is inferred that this dataset may be used for training or evaluating question answering systems, document retrieval, or natural language processing tasks, but the specific applications and background are not described in the README.

提供机构:
timchen0618
二维码
社区交流群
二维码
科研交流群
商业服务