遇见数据集

youssefkhalil320/pairs_three_scores_v5

收藏
Hugging Face2025-03-27 更新2025-08-30 收录
官方服务:

资源简介:

这是一个用于文本分类任务的英文数据集,包含约8000万条训练数据和约2000万条评估数据。每条数据由两个句子和一个表示两个句子相似度的分数组成。

This is an English dataset for text classification tasks, containing about 80 million training data and about 20 million evaluation data. Each data entry consists of two sentences and a score indicating the similarity between the two sentences.

提供机构:
youssefkhalil320
二维码
社区交流群
二维码
科研交流群
商业服务