相关数据集
Top 10 News
We scraped and parsed the homepages, politics pages, and top10 lists of prominent news sites for 2012 and 2016--2017. We did all this in 2016--2017, and hence the 2012 data exclusively comes from Inte
Mendeley Data2024-01-31 更新140
ComplexDataLab/cdl-data-chai-ccnews-20250513-dry-run
该数据集包含文章的标题、摘要、正文等字段,以及文章的作者、出版商、主题等信息。数据集被划分为训练集,可用于文本分类、信息抽取等自然语言处理任务。
Hugging Face2025-05-13 更新80
pnsahoo/newswire-1890-1899
该数据集包含新闻文章及其相关元数据,如作者、日期、报纸信息、主题标签、命名实体、地理位置等。数据集分为训练集,包含96740个样本。
Hugging Face2024-12-08 更新150
Swetha92/news_result_10
该数据集包含多个字段,包括ID、主题、摘要、批判性思维评分、沟通技能评分、自我意识评分以及元数据(如作者、日期和上下文)。数据集主要用于训练模型,可能涉及文本分析和评分预测等任务。数据集分为训练集,包含904个样本,总大小为4541035字节。
Hugging Face2024-12-13 更新120



