遇见数据集

Goekdeniz-Guelmez/sentence-compression-pairs

收藏
Hugging Face2025-11-13 更新2025-11-15 收录
官方服务:

资源简介:

该数据集包含三个部分:训练集、验证集和测试集。每个部分包含不同数量的文本对(anchor和positive),用于某种文本匹配或相似度度量的任务。数据集总共包含约180,000个示例,文件大小为36.78GB。

The dataset consists of three parts: training set, validation set, and test set. Each part contains a different number of text pairs (anchor and positive) for some text matching or similarity measurement task. The dataset includes a total of approximately 180,000 examples, with a file size of 36.78GB.

提供机构:
Goekdeniz-Guelmez
二维码
社区交流群
二维码
科研交流群
商业服务